Transparent pricing with no hidden fees. Pay as you go.
| Model | Spec | PoYo price | Official price | You save |
|---|---|---|---|---|
480p - with video input | Billed seconds = output duration + reference_video_urls duration | $0.058/sec 11.5 credits /sec | $0.091/sec Fal | 37% | |
480p / no video input | $0.100/sec 20 credits /sec | $0.151/sec Fal | 34% | |
720p - with video input | Billed seconds = output duration + reference_video_urls duration | $0.125/sec 25 credits /sec | $0.181/sec Fal | 31% | |
720p / no video input | $0.200/sec 40 credits /sec | $0.302/sec Fal | 34% | |
1080p - with video input | Billed seconds = output duration + reference_video_urls duration | $0.310/sec 62 credits /sec | $0.343/sec Fal | 9% | |
1080p / no video input | $0.450/sec 90 credits /sec | $0.685/sec Fal | 34% | |
4k - with video input | Billed seconds = output duration + reference_video_urls duration | $0.640/sec 128 credits /sec | $0.625/sec Fal | - | |
4k / no video input | $1.00/sec 200 credits /sec | $1.25/sec Fal | 20% | |
480p - with video input | Billed seconds = output duration + reference_video_urls duration | $0.045/sec 9 credits /sec | $0.073/sec Fal | 38% | |
480p / no video input | $0.070/sec 14 credits /sec | $0.121/sec Fal | 42% | |
720p - with video input | Billed seconds = output duration + reference_video_urls duration | $0.100/sec 20 credits /sec | $0.145/sec Fal | 31% | |
720p / no video input | $0.140/sec 28 credits /sec | $0.242/sec Fal | 42% |
* Actual fees are based on the final output.
Complete guide to using Seedance 2 API for Multimodal Video Generation
Seedance 2 API for Multimodal Video Generation
Available Seedance 2.0 API Models on PoYo
Seedance 2.0 Image to Video
Seedance 2.0 Reference to Video
Key Features of Seedance 2.0 API
ByteDance's director-level multimodal video model with unified audio-video generation, multi-shot storytelling, and cinematic motion control.
01
Unified Audio-Video Joint Generation
Seedance 2.0 uses a Dual-Branch Diffusion Transformer architecture to generate video and audio together in a single pass. Sound effects, dialogue, music, and ambient atmosphere are produced natively alongside the visuals, eliminating the need for separate sound design or post-production audio alignment.
- Native audio generated with video
- Dialogue, SFX, music, and ambient tracks
- Dual-Branch Diffusion Transformer
02
Multimodal Inputs: Text, Image, Video, and Audio
Seedance 2.0 accepts text prompts together with up to 9 images, 3 video clips (≤15s total), and 3 audio files (≤15s total) in a single request—up to 12 combined reference assets. This flexible input system lets creators blend visual references, motion cues, and audio signals to express complex creative intent in one generation.
- Up to 12 combined reference assets
- Mix text, image, video, and audio
- Structured reference tagging
03
Native Multi-Shot Storytelling
Generate cohesive multi-shot sequences from a single prompt. Seedance 2.0 maintains character consistency, visual style, and narrative continuity across camera cuts and transitions, making it ideal for short films, ads, and social content without manual stitching between clips.
- Multiple shots from one prompt
- Character and style consistency
- Seamless cuts and transitions
04
Native Dialogue and Lip-Sync in 8+ Languages
Seedance 2.0 produces clear, timing-accurate dialogue with lip-sync across more than eight languages, including English, Chinese, Japanese, and Korean. Speech, intonation, and facial motion are coupled at generation time, so localized content looks and sounds natural without post-processing alignment.
- Lip-sync across 8+ languages
- Clear, accurate dialogue timing
- Natural intonation and mouth motion
05
Director-Level Cinematic Control
Full creative influence over performance, lighting, shadows, camera movements, and physics. Seedance 2.0 understands directives like tracking, orbiting, and panning, delivers realistic motion stability, and aligns output with industry cinematic standards—making it excellent for action, dance, and fight sequences.
- Camera motion: tracking, orbit, pan
- Realistic physics and motion stability
- Cinematic lighting and shadow control
06
Video Extension, Style Transfer, and Enhancement
Beyond fresh generation, Seedance 2.0 supports conversational refinement, video extension, style transfer, and enhancement workflows. Use it to push existing footage further—extend scenes, restyle looks, or refine outputs iteratively within the same API surface.
- Extend existing video clips
- Style transfer and restyling
- Iterative refinement workflows
Who Can Benefit from Seedance 2.0 API?
Film and Pre-Visualization
Generate multi-shot pre-vis sequences with synchronized dialogue and ambient audio from a single prompt. Directors can test pacing, blocking, and camera work before committing to physical production.
Multilingual Content Localization
Adapt a single creative brief into regionally authentic video variants with native lip-sync across English, Chinese, Japanese, Korean, and more. Replace manual dubbing with model-driven localization.
Advertising and Marketing
Create cinematic ad concepts by combining product images, brand audio, reference footage, and copy in one request. Ideal for rapid iteration on hooks, reels, and social-first video campaigns.
Gaming and Interactive Media
Produce cutscenes, action sequences, and in-game video assets with consistent characters and immersive native soundtracks. Feed character art and audio cues to drive coherent narrative shots.
Education and Training
Build instructional video courses with consistent presenters and multilingual narration. Multi-shot storytelling keeps lessons visually engaging while native audio removes separate voiceover work.
Music Videos and Cinematic B-Roll
Leverage audio references to drive rhythm-synced visuals, or generate cinematic b-roll with strong motion stability. Seedance 2.0 excels at dance, action, and detailed gesture sequences.
Seedance 2.0 API vs Seedance 1.0, Sora 2, Veo 3.1 — Model Comparison
High-level comparison based on publicly available information as of April 2026. Specs may vary by provider, platform, and configuration.
| Feature | Seedance 2.0 | Seedance 1.0 | Sora 2 | Veo 3.1 |
|---|---|---|---|---|
| Architecture | Unified audio-video joint model | Video-focused diffusion | Video + audio model | Video + audio model |
| Input Modes | Text, Image, Video, Audio | Text, Image | Text, Image | Text, Image, First/Last Frames |
| Reference Assets | Up to 12 (9 img / 3 vid / 3 audio) | 1 image | 1 image | 2 frames |
| Typical Clip Length | Up to 15 seconds | Shorter, fewer shots | 15–25 seconds | 4–8 seconds |
| Max Resolution | Up to 2K | 1080p | Up to 1024p | Up to 4K |
| Native Audio | Yes, joint generation | Limited | Yes | Yes |
| Lip-Sync Languages | 8+ languages | Limited | Not publicly specified | Native dialogue |
| Multi-Shot Storytelling | Native, from one prompt | Manual stitching | Storyboard support | Improved narrative control |
| Video Extension | Yes | Limited | Limited | Limited |
Specifications evolve rapidly and depend on hosting platform and tier. Verify current details in the PoYo dashboard and documentation.
How to Use Seedance 2.0 API on PoYo
Step 1: Register & Create API Key
Sign up on PoYo and generate your API key instantly. No waitlist, no approvals required.
Create API Key →
Step 2: Top Up Credits
Add credits to your account. Pricing depends on resolution, duration, and selected mode (text-to-video, image-to-video, or multimodal reference).
Add Credits →
Step 3: Integrate the API
Submit a generation request with your prompt and reference files, then poll or webhook for results via the async workflow.
View Documentation →
Seedance 2.0 API Pricing: Pay Per Second
Seedance 2.0 API Pricing on PoYo (per second of video):
| seedance-2 · 1080p · text/image-to-video | 90 credits ($0.45/s) |
| seedance-2 · 1080p · with video input | 45 credits ($0.225/s) |
| seedance-2 · 720p · text/image-to-video | 40 credits ($0.20/s) |
| seedance-2 · 720p · with video input | 20 credits ($0.10/s) |
| seedance-2 · 480p · text/image-to-video | 20 credits ($0.10/s) |
| seedance-2 · 480p · with video input | 10 credits ($0.05/s) |
| seedance-2-fast · 720p · text/image-to-video | 28 credits ($0.14/s) |
| seedance-2-fast · 720p · with video input | 16 credits ($0.08/s) |
| seedance-2-fast · 480p · text/image-to-video | 14 credits ($0.07/s) |
| seedance-2-fast · 480p · with video input | 8 credits ($0.04/s) |
✓ Pay per second — no subscription fees
✓ Credits never expire
✓ seedance-2-fast is ~30% cheaper for draft iteration
✓ 1080p is available on seedance-2 only
✓ Video input mode unlocks the lowest tier
Start from as little as $0.04 per second. Top up credits →
Frequently Asked Questions about Seedance 2.0 API
What is Seedance 2.0?
What inputs does Seedance 2.0 support?
What are the video length and resolution limits?
Does Seedance 2.0 support native audio and lip-sync?
How does Seedance 2.0 compare to Seedance 1.0?
Can I use Seedance 2.0 for commercial projects?
How quickly can I start using Seedance 2.0 API?
Why Use PoYo for Seedance 2.0 API Access
Instant Access
Create your API key and start generating in minutes. No waitlist, no regional restrictions, no extra approvals.
Affordable Credit-Based Pricing
Pay only for what you generate with transparent credits. No subscriptions, and credits never expire.
Developer-Friendly Async API
Submit jobs and fetch results with a clean async workflow. Easy integration into any stack with clear docs.
99.9% Uptime and 24/7 Support
Production-ready infrastructure with reliable uptime, stable access, and responsive technical support.
Free Playground Access
Test prompts and references on the model page before integrating—no credit card required to explore.
Unified Model Library
Access Seedance 2.0 alongside other leading AI video and image models through one platform and one billing system.
