Advanced video generation with Seedance 2.0
Seedance 2.0
This video captures a man in vintage aviator attire, including a leather helmet and goggles, actively piloting a unique, bicycle-like flying machine. The overall mood is adventurous and slightly whimsical, set against a bright outdoor backdrop of rolling green hills under a clear blue sky. The movement is dynamic, with the propeller blurred to suggest motion, and the man appears focused and determined. The visual style is realistic, almost cinematic, with a warm color palette dominated by earthy browns, greens, and blues, and high saturation. The subject is an inventive individual operating a quirky, self-made aircraft, evoking a sense of historical innovation or steampunk fantasy.
Seedance 2.0
This video captures a man in vintage aviator attire, including a leather helmet and goggles, actively piloting a unique, bicycle-like flying machine. The overall mood is adventurous and slightly whimsical, set against a bright outdoor backdrop of rolling green hills under a clear blue sky. The movement is dynamic, with the propeller blurred to suggest motion, and the man appears focused and determined. The visual style is realistic, almost cinematic, with a warm color palette dominated by earthy browns, greens, and blues, and high saturation. The subject is an inventive individual operating a quirky, self-made aircraft, evoking a sense of historical innovation or steampunk fantasy.
Multimodal video generation with synchronized audio
Generate videos from text, images, video clips, and audio references combined. Seedance 2.0 produces video and sound together in a single pass, with no post-production sync required.
Multimodal input
Combine text, up to 9 reference images, 3 video clips, and 3 audio clips in a single generation for precise creative control.
Audio-visual joint generation
Video and audio generated in one pass. Dialogue, music, ambient sound, and foley are synchronized from the start. Dual-channel stereo.
Reference-driven control
Supply reference images, videos, and audio to anchor visual style, camera movement, and pacing. The model preserves these across generation.
Complex motion and physics
Realistic multi-subject interactions, sports footage, crowd scenes, and choreography with physically plausible motion and detail.
Video editing and extension
Regenerate specific sections with new prompts and extend video length with continuous motion and consistent subjects.
Full production stack
Connect to Text to Speech, lip-sync, Eleven Music, AI Sound Effects, and Flows for end-to-end video production in one platform.
Create with Seedance 2.0 with full control
Create with Seedance 2.0 with full control
Select Seedance 2.0
Select Seedance 2.0 from the ElevenCreative model shelf to start generating video with synchronized audio.
Enter prompt and add references
Describe your scene with a text prompt and optionally add reference images, video clips, or audio to guide style and motion.
Generate and refine
Generate your video, then refine sections, extend length, or add narration, music, and sound effects in Studio.
Frequently asked questions
Get started with Seedance 2.0 for free
The best image, video, and audio models — all inside ElevenCreative. Start generating today.

Discover more image & video generation models
Explore our full library of AI image and video generation models, each with unique strengths and capabilities.

Seedream 5.0 Pro creates controllable commercial images, dense layouts, and realistic product visuals.

Wan 2.5 Video generates short, high-resolution videos from text or images with smooth motion and flexible control.

Seedream 5 Lite rapidly generates images up to 3K with deep reasoning, precise prompt-following

HeyGen Avatar 4 quickly generates expressive, lifelike talking avatars from a single photo

Creatify Aurora transforms a single photo and audio into ultra-realistic AI avatar videos with nuanced motion and visual style

Veed Fabric 1.0 creates realistic talking videos from any image, with lifelike gestures, lip sync, and flexible style control.

Sync 3 offers 4K lip synced video generation with shot-level realism, emotion preservation

Gemini Omni Flash creates realistic, high-resolution video from text, images, and video prompts all on ElevenLabs.
