
AI API untuk pengembang
API Teks ke Video
Bangun pembuatan video AI ke dalam aplikasi Anda dengan PoYo API Teks ke Video. Bandingkan model video, durasi, resolusi, opsi audio, dan harga untuk alur kerja materi iklan otomatis.
Model yang cocok
33 model

$0.105 /detik
hailuo-03
MiniMax H3 (Hailuo 03) generates native 2K video from text, first/last frames, or image, video, and audio references.

$0.085 /detik yang ditagih
seedance-2.5
ByteDance Seedance 2.5 video generation with 4-30 second output, first and last frame control, up to 50 image, video, and audio references, optional synchronized audio, and 480p, 720p, or 1080p delivery.

$0.17 /detik
flux-3/text-to-video
Black Forest Labs' FLUX 3 video family for text, images, first and last frames, source-video extension, positioned keyframes, and native audio generation.

$0.045 /detik video keluaran dan referensi
seedance-2
ByteDance Seedance 2 video generation with text-to-video, image-to-video, multimodal reference-to-video, video-to-video, and audio-to-video workflows using image, video, and audio references with optional native audio.

$0.225 /generasi
gemini-omni-1.1-flash
Gemini Omni 1.1 Flash API for text-to-video, first and last frame interpolation, reference image generation, and reference-video workflows.

$0.03 /detik video keluaran dan referensi
seedance-2-mini
ByteDance Seedance 2.0 Mini video generation for lower-cost text-to-video, image-to-video, and multimodal reference workflows with 480p and 720p output.

$0.11 /detik
happy-horse-1.1
Alibaba Happy Horse 1.1 video generation with text-to-video, image-to-video, and reference-to-video workflows.

$0.135 /detik
kling-3.0/standard
Kuaishou's advanced Kling 3.0 video model family with Standard, Pro, and native 4K generation, multi-shot storyboarding, synchronized audio, and element-based consistency.

$0.05 /detik
kling-o3/standard
Kling O3 video generation on PoYo with Standard, Pro, and 4K variants for text-to-video, image-to-video, and reference-to-video workflows.

$0.20 /video
sora-2-official
OpenAI Sora 2 on PoYo with synced audio, improved physics, optional reference image input, and fixed 4-second, 8-second, 12-second, 16-second, and 20-second tiers.

$0.06 /detik
wan2.7-text-to-video
Wan 2.7 Video models on PoYo for text-to-video, image-to-video, reference-to-video, and edit-video workflows.

$0.05 /detik
wan3.0-text-to-video
Generate videos up to 30 seconds with Wan 3.0 using text, first and last frames, or multimodal references including images, video, audio, documents, and links.

$0.068 /detik
wan3.0-prime-text-to-video
Wan 3.0 Video Prime is Alibaba's Wan 3.0 tier optimized for faster turnaround, with multimodal image, video, audio, document, and public webpage references, videos up to 30 seconds, and optional generated audio.

Akan segera datang.
Kling 4.0
Kling 4.0 brings 30-second cinematic video, 4K output, multi-keyframe control, multimodal references, and native stereo audio to creative workflows.

Akan segera datang.
Kling 4.0 Flash
Kling 4.0 Flash brings fast 720p video creation up to 20 seconds, native audio, and Omni Reference for text, image, and multimodal creative workflows.

$0.05 /detik
h3-max
MiniMax H3 Max generates 5–15 second videos from text, first/last frames, or image, video and audio references at 480p, 768p and 1080p.

$0.025 /detik
h3-max-turbo
MiniMax H3 Max Turbo generates 5–15 second videos from text or first/last frames at 480p, 768p and 1080p.

$0.035 /detik
hailuo-02
MiniMax's #2 globally-ranked video model with NCR architecture, ultra-realistic physics, and 1080p cinematic output.

$0.375 /video
runway-gen-4.5
Runway Gen-4.5 is a high-fidelity video model focused on prompt adherence, cinematic motion, visual fidelity, and optional reference image guidance.

$0.10 /video
veo3.1-lite
Google DeepMind's VEO 3.1 video family with text-to-video, image-to-video, 720p, 1080p, and 4K output options.

$0.018 /detik
veo3.1-fast-official
Official VEO 3.1 video generation with text, image, first/last-frame and reference workflows across fast, lite and quality tiers.

$0.03 /generasi
wan2.2-image-to-video-fast
Wan 2.2 Fast provides fast text-to-video and image-to-video generation with low-cost 480p and 720p tiers for quick iteration.

$0.175 /video
hailuo-2.3
MiniMax's Hailuo 2.3 video model for realistic human motion, expressive characters, and text-to-video or first-frame guided generation at 768p and 1080p.

$0.60 /generasi
omni-flash
Omni Flash API access for text-to-video, image-to-video, three-image reference fusion, and video-input generation workflows.

$0.105 /video
seedance-1.0-pro
ByteDance's #1 ranked video model with multi-shot storytelling, cinema-grade motion, and bilingual text-to-video generation.

$0.045 /video
seedance-1.5-pro
ByteDance's latest video model with synchronized audio generation, flexible aspect ratios, and enhanced motion control.

$0.15 /video
wan2.5-image-to-video
Wan 2.5 combines text-to-video and image-to-video generation with 5-second and 10-second output, synchronized audio support, and multiple size and resolution tiers.

$0.045 /detik
kling-1.6/standard
Kling 1.6 Standard and Pro video generation on PoYo with text, first/last frame, and multi-image element reference workflows.

$0.325 /video
kling-2.6
Kuaishou's revolutionary video model that simultaneously generates visuals with synchronized dialogue, sound effects, and ambient audio in one pass.

$0.40 /video
wan2.6-text-to-video
Alibaba's Wan 2.6 video generation family for text-to-video, image-to-video, and video-to-video with multi-shot 1080p output.

$0.21 /video
kling-2.5-turbo-pro
Kling 2.5 Turbo Pro is a flexible short-form video model with text-to-video, optional frame guidance, smooth motion, cinematic depth, and fixed 5-second and 10-second tiers.

$0.085 /detik
kling-3.0-turbo/standard
Kling 3.0 Turbo is the speed-optimized Kling 3.0 variant for high-volume text-to-video and image-to-video generation with multi-shot storyboarding. Standard (720p) and Pro (1080p) models available.

$0.08 /detik
happy-horse
Alibaba Happy Horse 1.0 video generation and editing with text-to-video, image-to-video, reference-to-video, and video-edit workflows.
Teks ke video Model API - Harga dan Model yang Cocok
Bandingkan penyedia, masukan yang didukung, format keluaran dan harga sebelum memilih model.
Cocok dengan tugas
Model muncul di sini hanya jika katalog metadata mereka termasuk Text to Video.
Perbandingan penyedia
Ulasan keluarga model di seluruh penyedia tanpa meninggalkan direktori tugas.
API-jalan siap
Setiap kartu model menghubungkan ke PoYo harga, contoh, kontrol taman bermain dan API details.
| Model | Penyedia | Jenis tugas | Harga |
|---|---|---|---|
| MiniMax H3 / Hailuo 03 hailuo-03 | MiniMax | Text to Video, Image to Video | $0.105/detik |
| Seedance 2.5 seedance-2.5 | Seedance | Text to Video, Image to Video, Video to Video | $0.085/detik yang ditagih |
| FLUX 3 flux-3/text-to-video | Black Forest Labs | Text to Video, Image to Video | $0.17/detik |
| Seedance 2 seedance-2 | Seedance | Text to Video, Image to Video, Video to Video | $0.045/detik video keluaran dan referensi |
| Gemini Omni 1.1 Flash gemini-omni-1.1-flash | Text to Video, Image to Video, Video to Video | $0.225/generasi | |
| Seedance 2.0 Mini seedance-2-mini | Seedance | Text to Video, Image to Video, Video to Video | $0.03/detik video keluaran dan referensi |
| Happy Horse 1.1 happy-horse-1.1 | Alibaba | Text to Video, Image to Video, Uncensored | $0.11/detik |
| Kling 3.0 kling-3.0/standard | Kling | Text to Video, Image to Video | $0.135/detik |
| Kling O3 kling-o3/standard | Kling | Text to Video, Image to Video | $0.05/detik |
| Sora 2 Official sora-2-official | OpenAI | Text to Video, Image to Video | $0.20/video |
| Wan 2.7 Video wan2.7-text-to-video | Alibaba | Text to Video, Image to Video, Uncensored | $0.06/detik |
| Wan 3.0 wan3.0-text-to-video | Alibaba | Text to Video, Image to Video, Uncensored | $0.05/detik |
| Wan 3.0 Video Prime wan3.0-prime-text-to-video | Alibaba | Text to Video, Image to Video, Uncensored | $0.068/detik |
| Kling 4.0 kling-4-0 | Kling | Text to Video, Image to Video | Akan segera datang. |
| Kling 4.0 Flash kling-4-0-flash | Kling | Text to Video, Image to Video | Akan segera datang. |
| MiniMax H3 Max h3-max | MiniMax | Text to Video, Image to Video | $0.05/detik |
| MiniMax H3 Max Turbo h3-max-turbo | MiniMax | Text to Video, Image to Video | $0.025/detik |
| Hailuo 02 hailuo-02 | MiniMax | Text to Video, Image to Video | $0.035/detik |
| Runway Gen-4.5 runway-gen-4.5 | Runway | Text to Video | $0.375/video |
| Veo 3.1 veo3.1-lite | Text to Video, Image to Video | $0.10/video | |
| Veo 3.1 Official veo3.1-fast-official | Text to Video, Image to Video | $0.018/detik | |
| Wan 2.2 Fast wan2.2-image-to-video-fast | Alibaba | Text to Video, Image to Video | $0.03/generasi |
| Hailuo 2.3 hailuo-2.3 | MiniMax | Text to Video, Image to Video | $0.175/video |
| Omni Flash omni-flash | Text to Video, Image to Video, Video to Video | $0.60/generasi | |
| Seedance 1.0 Pro seedance-1.0-pro | Seedance | Text to Video, Image to Video | $0.105/video |
| Seedance 1.5 Pro seedance-1.5-pro | Seedance | Text to Video, Image to Video | $0.045/video |
| Wan 2.5 wan2.5-image-to-video | Alibaba | Text to Video, Image to Video | $0.15/video |
| Kling 1.6 kling-1.6/standard | Kling | Text to Video, Image to Video | $0.045/detik |
| Kling 2.6 kling-2.6 | Kling | Text to Video, Image to Video | $0.325/video |
| Wan 2.6 wan2.6-text-to-video | Alibaba | Text to Video, Image to Video, Video to Video | $0.40/video |
| Kling 2.5 Turbo Pro kling-2.5-turbo-pro | Kling | Text to Video, Image to Video | $0.21/video |
| Kling 3.0 Turbo kling-3.0-turbo/standard | Kling | Text to Video, Image to Video | $0.085/detik |
| Happy Horse happy-horse | Alibaba | Text to Video, Image to Video, Uncensored | $0.08/detik |
Pertanyaan yang sering diajukan
Apa itu API Teks ke Video?
+
API teks ke video mengubah deskripsi tertulis menjadi klip yang dihasilkan. Aplikasi Anda menentukan adegan dan pengaturan yang didukung, lalu mengambil video setelah diproses. Ini cocok untuk pembuatan gambar, pembuatan prototipe kreatif, dan produk yang menawarkan pembuatan video AI kepada penggunanya.
Bagaimana cara menulis perintah untuk API pembuatan video?
+
Jelaskan satu adegan dengan jelas: subjek, aksi, latar, pergerakan kamera, dan gaya visual. Gunakan kata-kata gerak yang dapat terjadi dalam durasi yang dipilih. Hindari mengemas keseluruhan cerita ke dalam satu klip pendek; pisahkan pengambilan gambar yang berbeda menjadi permintaan yang dapat disusun oleh aplikasi Anda nanti.
Model teks ke video manakah yang terbaik untuk aplikasi saya?
+
Bandingkan model berdasarkan hasil jepretan yang representatif daripada hanya mengandalkan satu video etalase. Evaluasi gerakan, stabilitas visual, akurasi cepat, audio dan waktu pemrosesan serta harga. Jaga agar pengaturan masukan tetap sama dan catat berapa banyak upaya yang menghasilkan klip yang dapat digunakan untuk aplikasi spesifik Anda.
Bisakah API teks ke video menghasilkan suara dan dialog?
+
Beberapa model video menghasilkan audio asli, sementara model lainnya menampilkan klip senyap atau menyediakan kontrol audio opsional. Periksa titik akhir spesifik untuk perilaku ucapan, musik, dan efek suara. Jika narasi yang tepat diperlukan, pertimbangkan untuk membuat sulih suara terpisah melalui Text to Speech dan menggabungkannya dalam alur kerja media Anda.
Berapa durasi video yang dihasilkan?
+
Opsi durasi bergantung pada model dan titik akhir. Pilih durasi yang terdokumentasi, bukan angka sembarangan. Untuk konten yang lebih panjang, hasilkan gambar individual dan susun, atau gunakan fitur ekstensi yang didukung oleh model yang dipilih. API klip pendek tidak secara otomatis membuat film yang telah diedit secara lengkap.
Resolusi dan rasio aspek manakah yang harus saya pilih?
+
Pilih resolusi dan rasio aspek yang tersedia yang sesuai dengan penempatan akhir. Dimensi yang lebih tinggi dapat meningkatkan biaya atau waktu pemrosesan tanpa memperbaiki gerakan yang lemah atau konten yang tidak akurat. Uji konfigurasi draf praktis terlebih dahulu, lalu konfirmasikan pengaturan keluaran akhir untuk model yang dipilih.
Bagaimana cara meningkatkan konsistensi di beberapa klip?
+
Gunakan kembali deskripsi subjek dan gaya yang jelas, rencanakan pengambilan gambar sebelum pembuatan, dan hindari mengubah terlalu banyak variabel di antara permintaan. Jika identitas atau tampilan harus mengikuti desain tetap, pindah ke input referensi gambar atau multimodal yang didukung. Klip teks saja yang dibuat secara independen mungkin berbeda dalam karakter, objek, atau pemandangan.
Apakah teks ke video sama dengan merender skrip dengan rekaman stok?
+
Tidak. Model video generatif mensintesis klip baru dari instruksi. Render template, perakitan subtitle, dan pengeditan rekaman stok adalah alur kerja terpisah yang dapat digabungkan dengan klip yang dihasilkan. Gunakan dokumentasi tugas dan model PoYo untuk memilih kemampuan pembangkitan yang sesuai dengan saluran Anda.
Bagaimana cara mengelola permintaan API teks ke video?
+
Kirimkan model dan input yang didukung ke /api/generate/submit dan simpan task_id yang dikembalikan. Kueri /api/generate/status/{task_id} atau gunakan callback_url untuk menerima peristiwa penyelesaian. Pertahankan status pekerjaan dalam aplikasi Anda sehingga penyegaran browser atau waktu tunggu klien tidak kehilangan jejak permintaan video yang sedang berjalan.
Bagaimana cara menghitung harga API teks ke video?
+
Penetapan harga video bisa per generasi, per klip, atau per detik, dengan tingkatan durasi, resolusi, kualitas, dan audio. Bandingkan pengaturan aktual dan unit penagihan untuk setiap model. Sertakan biaya draf materi iklan yang ditolak saat memperkirakan pengeluaran yang diperlukan untuk produksi akhir.
Mengapa video saya masih diproses atau mengapa gagal?
+
Pembuatan video bisa memakan waktu lebih lama dibandingkan permintaan HTTP normal. Periksa status tugas yang disimpan sebelum mencoba lagi dan periksa kesalahan yang muncul untuk mengetahui pengaturan yang tidak didukung, masukan yang hilang, atau masalah akun. Jangan perlakukan waktu tunggu klien sebagai bukti kegagalan atau membuat pekerjaan duplikat saat tugas awal sedang berjalan.
Di mana saya dapat menemukan contoh kode API video?
+
Buka halaman model yang dipilih untuk dokumentasi dan contoh API saat ini, atau baca /models/{model-id}/llms.txt. Gunakan bidang tersebut untuk klien Python, JavaScript, atau HTTP Anda. Simpan kredensial di server dan pisahkan pekerjaan pembuatan klip dari antarmuka yang menampilkan kemajuannya.
Menjelajahi semua model AI
Buka pasar model penuh untuk membandingkan gambar, video, audio, chat, 3D dan model alat.
Lihat semua modelMembangun dengan API
Buka halaman model untuk parameter, harga dan dokumentasi API.
Dokumentasi APIHalaman tugas yang terkait