New Model

Seedance 2.0

ByteDance's multimodal AI video generator supports text, image, audio, and video inputs, with model-dependent controls for duration, audio, and reference files.

Why Seedance 2.0?

A unified multimodal architecture that generates audio and video together, with unprecedented creative control over AI video generation.

Multimodal Input

Seedance 2.0 supports text, image, audio, and video as input — four modalities for AI video generation. Upload up to 9 images, 3 videos, and 3 audio files to guide generation with pixel-level precision.

Character Consistency

Use references and detailed prompts to guide characters and scenes across longer AI-generated shots. Results vary by scene and input.

Dual-Channel Audio

Seedance 2.0 generates synchronized dual-channel stereo audio — dialogue, sound effects, and background music — natively aligned with your AI-generated video content.

Smart Editing & Extension

Edit existing AI-generated videos by replacing characters, modifying scenes, or extending clips. Seamlessly continue your story with stable visual continuity using Seedance 2.0's targeted editing.

Physics-Accurate Motion

Seedance 2.0 is designed to handle multi-subject movement, gravity, fluids, and material behavior. Results vary with the prompt and source media.

Director-Level References

Reference composition, camera movement, motion rhythm, and visual effects from your input assets in Seedance 2.0. Upload storyboards, reference videos, and audio clips — your references become your creative direction for AI video generation.

See What Seedance 2.0 Can Create

From cinematic storytelling to product commercials — real AI video generation examples powered by Seedance 2.0.

Image to Video

Product Commercial

The character in the painting looks guilty, darting glances left and right before leaning out of the picture frame. Very swiftly, they extend a hand outside the frame, grab a cola, take a sip, and reveal a look of deep satisfaction. Hearing footsteps approach, the character hastily puts the cola back just as a Western cowboy walks by and picks it up. The ending shot pushes in on a top-lit close-up of the cola against a pure black background, with artistic subtitles and a voiceover appearing at the bottom: "Yikou Cola, an absolute must-try!"

Reference to Video

Extend the Video Length

Extend the video length. The camera follows a man in orange riding a brown horse. He speeds up and gallops to a large tree with orange flowers ahead, breaks off two blossoms from the branch. Then others ride into the frame one after another. The camera pushes in to shoot the man in orange clothes dismounting from the horse, then the camera quickly circles around him. He turns to walk toward a woman in white riding a white horse and presents the flowers to her. Style: Chinese classical lady painting, 3D, cheerful folk music, shadow puppet style, with black, white, and orange as the main color palette.

Reference to Video

Multimodal All-Round Reference

Refer to the shooting script in @Image 1, and draw on the storyboard, shot scale, camera movement, visuals and copy in @Image 1. The character is from @Image 2, the scene is from @Image 3, and the props are from @Image 4. Create a 15-second healing short film.

Text to Video

Wuxia-Style Audiovisual Blockbuster

A Wuxia-style audiovisual blockbuster. A white-clad swordsman and a straw-caped blademaster face off in a bamboo forest. The camera slowly pushes in between them, shifting focus between raindrops and sword hilts. The atmosphere is extremely oppressive; only the sound of rain can be heard. Suddenly, a crack of thunder flashes, and both charge simultaneously. A fast-panning profile shot captures their mud-splattering footsteps. The precise moment their weapons clash, the footage switches to ultra-slow motion, clearly displaying the ring-shaped shockwaves of rainwater blasted away by the blades, along with bamboo leaves sliced by sword aura. Normal speed resumes as they land back-to-back. The straw-caped blademaster's bamboo hat splits open, and the scene cuts abruptly.

Text to Video

1920s Jazz Club Charleston Dance

1920s jazz club-style Charleston dance. A female dancer in a gold fringe dress and a male dancer in a striped suit perform a high-intensity routine. Moves include hyper-fast syncopated footwork, aerial catches, and exaggerated arm swings. The camera uses dynamic tracking, interspersed with close-ups of footwork. Emphasis is on physical details the fringe swinging wildly with every kick, the sheen of sweat on their skin, and the smoky, vintage film grain cinematic texture. A background jazz band and cheering audience amplify the frenzied party atmosphere.

Image to Video

Girl Hangs Laundry Gracefully

A girl hangs laundry gracefully. After finishing, she takes another piece of clothing from the bucket and shakes it vigorously.

Key Capabilities of Seedance 2.0

Explore Seedance 2.0 video generation inputs, controls, and supported workflows

Seedance 2.0 Exclusive

4-Modality AI Video Generation

Unlike Seedance 1.5 Pro which supports only text and image input, Seedance 2.0 accepts four input modalities — text prompts, images, audio files, and video clips. Upload up to 15 reference files simultaneously for precise control over your AI-generated video.

Seedance 2.0 Exclusive

Native Audio-Visual Joint Generation

Seedance 2.0 can generate synchronized audio with supported outputs, including dialogue, sound effects, and background audio when the selected mode provides those controls.

Seedance 2.0 Exclusive

Targeted Video Editing

Seedance 2.0 introduces targeted editing for AI-generated videos — replace specific characters, modify individual scenes, or extend clips without regenerating the entire video. Seedance 1.5 Pro requires full regeneration for any change.

Seedance 2.0 Exclusive

Storyboard-to-Video Generation

Upload a storyboard or shooting script as a reference image and Seedance 2.0 will follow the shot scale, camera movement, and visual composition. The AI video generator transforms your creative direction into a polished video automatically.

Seedance 2.0 Advantage

Motion and Scene Dynamics

Seedance 2.0 is designed for scenes involving gravity, fluid motion, cloth, and multiple subjects. Output quality depends on the prompt, references, and selected settings.

Multi-shot Support

Character & Scene Consistency

Use references and detailed prompts to guide character identity, wardrobe, and environment across multiple shots. Consistency can vary by scene and source material.

Technical Specifications

Seedance 2.0 delivers professional-grade AI video generation with flexible configuration options.

Max Duration

15 seconds

Resolution

480P / 720P / 1080P

Aspect Ratios

16:9, 9:16, 1:1, 4:3, 3:4, 21:9

Audio Output

Dual-channel stereo (native)

Reference Files

Up to 15 (9 images + 3 videos + 3 audio)

Input Modalities

Text + Image + Audio + Video

Provider

ByteDance via APImart

Seedance 2.0 vs Seedance 1.5 Pro

Compare supported inputs and output controls with the previous model generation.

FeatureSeedance 2.0Seedance 1.5 Pro
Max Duration15s12s
Max Resolution1080P1080P
Audio OutputDual-channel StereoMono (coordinated)
Reference Files15 (9 img + 3 vid + 3 aud)3
Input Modalities4 (Text + Image + Audio + Video)2 (Text + Image)
Character ConsistencyYesGood
Physics AccuracyYesBaseline
Video ExtensionYesNo
Targeted EditingYesNo
Storyboard-to-VideoYesNo

Ready to Create with Seedance 2.0?

Start creating with Seedance 2.0 using the models and output options currently available.

Seedance 2.0 AI Video Generator | Turing AI