Skip to content
Dashboard

Seedance 2.0

Seedance 2.0 generates video with synchronized multilingual audio, professional camera work, multi-shot composition, and in-video text rendering. Inputs include text, image, multimodal reference, and existing video for editing and extension. Your use is subject to ByteDance's Terms & Privacy Policies.

Video GenVision (Image)image-to-videotext-to-videoreference-to-video
import { experimental_generateVideo as generateVideo } from 'ai';
const result = await generateVideo({
model: 'bytedance/seedance-2.0',
prompt: 'A serene mountain lake at sunrise.'
});
Read docs

Copy link to headingPlayground

Try out Seedance 2.0 by ByteDance. Usage is billed to your team at API rates. Free users (those who haven't made a payment) get $5 of credits every 30 days.

bytedance logo
Images(optional)
Add up to 9 images
Videos(optional)
Add up to 3 videos
Prompt(optional)

End frame(optional)
Duration8s
4s15s
Resolution
Aspect ratio
Videos to generate
bytedance logo

Your generated video will appear here.

Copy link to headingProviders

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Provider
Input
Output
Capabilities
ZDR
No Training
Release Date
$7/M+1 more
04/14/2026

Copy link to headingMore models by ByteDance

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Release Date
$10.70/M+1 more
bytedance logo
08/07/2026
$5.60/M+1 more
bytedance logo
04/14/2026
$0.01/sec+2 more
bytedance logo
10/24/2025
256K
1.1s
96tps
$0.25/M+1 more
$2/M+1 more
Read:
$0.05/M
Write:
bytedance logo
09/01/2025
256K
1.1s
79tps
$0.25/M+1 more
$2/M+1 more
Read:
$0.05/M
Write:
bytedance logo
09/01/2025
$0.02/sec+2 more
bytedance logo
06/11/2025

Copy link to headingAbout Seedance 2.0

Seedance 2.0 was released April 14, 2026 as the second-generation ByteDance Seedance video model. The standard variant targets the highest output quality in the 2.0 lineup.

Input modes span text-to-video, image-to-video, multimodal reference-to-video (combining image, video, and audio references), and video editing and extension. One model covers the full range of creative workflows where Seedance 1.0 Pro required separate variants.

Output supports 16:9 aspect ratio with 720p resolution in examples shown, and clip durations from five to 10 seconds. Seedance 2.0 maintains motion stability and fine detail across frames, handles complex scenes with facial expressions and physical interactions, and renders text inside generated video.

Native audio generation with multilingual support lets you produce dialogue, sound effects, and ambient audio without a separate text-to-speech or audio compositing step. Professional camera movements and multi-shot composition extend creative control beyond single static shots.

AI Gateway applies no markup on video generation for Seedance 2.0: the rate matches the direct ByteDance provider price. Set the model to bytedance/seedance-2.0 and call it through the AI SDK's generateVideo function.

Copy link to headingWhat To Consider When Choosing a Provider

  • Configuration: Seedance 2.0 accepts text, image, multimodal reference (image plus video plus audio), and existing video as input. If your pipeline handles multiple input modes, route through a single model rather than composing separate generation steps.
  • Zero Data Retention: AI Gateway does not currently support Zero Data Retention for this model. See the documentation for models that support ZDR.
  • Authentication: AI Gateway authenticates requests using an API key or OIDC token. You do not need to manage provider credentials directly.

Copy link to headingWhen to Use Seedance 2.0

Best for

  • Multimodal video workflows: Text, image, video, and audio inputs combined in a single reference-to-video generation
  • Character-driven content: Scenes with facial expressions, physical interactions, and synchronized dialogue in multiple languages
  • Cinematic production: Professional camera movements and multi-shot composition that extend beyond typical social clip defaults
  • In-video text rendering: Content where legible text inside the generated video matters for brand or narrative
  • Video editing and extension: Modifying existing video or extending a source clip without regenerating from scratch

Consider alternatives when

  • Maximum generation speed: Seedance 2.0 Fast trades some quality for faster turnaround and lower cost
  • Static image generation: Use a dedicated image model when motion isn't required
  • Video understanding only: Use a vision-language model when you need to analyze existing video rather than generate new content

Seedance 2.0 consolidates second-generation Seedance capabilities into a single model: multimodal inputs, high-fidelity motion, native synchronized audio, professional camera work, and in-video text. For teams producing character-driven or cinematic short-form video, it's the quality-focused default in the 2.0 line.