New: the Nano Banana 2 Lite model is live on Nano Banana.

See the model
ModelVideoComing soon

Seedance 2.0

ByteDance's director-level AI video: text, image, and audio inputs with cinematic stability

Seedance 2.0 is ByteDance's unified multimodal video model — it accepts text, image, and audio simultaneously and delivers cinematic, high-stability video with director-level control over lighting, shadows, and camera movement.

All models

Coming soon — not connected yet

We have not integrated this model. Everything on this page is documented from its published capabilities on our legacy site; nothing here generates output. Live generation on Nano Banana currently runs Nano Banana 2 Lite.

Use the live workbench

Capabilities

Why this model

  1. 01

    Unified Multimodal Synthesis

    Beyond simple prompts: text, image, and audio inputs work simultaneously for complex, multi-layered visual storytelling with perfect coherence.

  2. 02

    Director-Level Motion Control

    Control character performance, dynamic lighting, and intricate camera paths to produce professional, industry-standard cinematic visuals.

  3. 03

    Immersive Audio-Video Sync

    Joint audio-visual generation keeps motion and sound perfectly synchronized for a truly immersive, production-ready experience.

Best for

Where it shines

  • Professional film and marketing workflows
  • Image-to-video with preserved detail
  • Audio-synchronized cinematic scenes
  • Commercial social content

Specs

Technical snapshot

Resolution
Coming soon
Aspect ratio
Coming soon
Speed
Coming soon
Pricing tier
Coming soon

Specs are documented from the model's legacy page; fields marked Coming soon were not published. They describe the model itself — availability on Nano Banana is shown by the status badge.

FAQ

Common questions

ByteDance's latest AI video model, built on a unified multimodal architecture that accepts text, image, and audio inputs for superior video generation.