Skip to main content
Model logo

SeeDance 2.5 API

  • Text to Video
  • Image to Video
  • Reference to Video
Generate Audio
NSFW Check
Output

Your generated video will appear here

SeeDance 2.5

What is SeeDance 2.5?

SeeDance 2.5 is ByteDance's August 2026 audio-video model for longer stories, precise reference control, and video editing. It is live on Unifically as bytedance/seedance-2.5. The callable route generates 4 to 30 second videos at 480p, 720p, or 1080p, with synchronized audio on by default.

Three workflows share the same model ID: text-to-video, first-and-last-frame generation, and multimodal reference generation. Reference mode accepts images, videos, and audio in one request. This is the biggest practical change from the older SeeDance line: one generation can carry much more source material across a longer timeline.

What's new in SeeDance 2.5

  • Up to 30 seconds. Duration is any whole second from 4 through 30, twice the 15-second ceiling on SeeDance 2.0.
  • A 50-asset reference budget. Send up to 30 images, 10 videos, and 10 audio files, capped at 50 assets total.
  • Longer reference media. Reference videos can total 30 seconds. Reference audio can also total 30 seconds.
  • Audio-only reference runs. A prompt plus audio references is valid; an image or video reference is not required.
  • First and last frame control. Pin both ends of the video while the model creates the motion between them.
  • Native audio by default. Voice, effects, and background music generate with the picture in the same run.

Best for

Longer narrative videos

A 30-second ceiling gives a short story room for multiple beats without joining separate generations.

Reference-heavy campaigns

Combine product images, motion references, voice, and music cues in one request.

Pinned opening and closing shots

Set the first and last frame, then generate the transition between them.

Dialogue and sound-led scenes

Native audio is on by default, and audio-only reference requests are supported.

Storyboard-driven production

Use numbered image, video, and audio references directly inside the prompt.

Use cases

Build a 30-second product story that moves from a wide establishing shot to a close product reveal without joining separate videos. Feed a campaign's product angles, style frames, camera reference, and music cue into one generation. Pin a first frame and final call-to-action frame for an ad with a fixed open and close. For character work, use reference images plus a motion video and address each asset as [Image1] or [Video1] in the prompt. Audio-led workflows can start from a voice or rhythm reference without adding an image.

Limitations

The current API route exposes 480p, 720p, and 1080p. The 4K output and region-editing tools shown in ByteDance's consumer product are not callable parameters here. First-and-last-frame mode requires both images and cannot be mixed with reference arrays. Reference videos and audio each have a 30-second combined limit. Arena has not added Seedance 2.5 to its video boards yet, so there is no independent rank or Elo score for this version as of August 7, 2026.

Draft mode: 480p first, 1080p when you like it

bytedance/seedance-2.5-draft generates a 480p draft so you can test prompts and references before paying for a full-resolution render. When a draft is right, send its task ID as draft_task_id. The same video is then generated natively in 1080p from that task, without any external upscaler. Prompt, duration, aspect ratio, and references all come from the draft, so nothing needs to be re-entered. In the playground, choose Draft and paste the finished draft's task ID into Draft Task ID. 720p and 4K renders from a draft are coming soon.

SeeDance 2.5 vs SeeDance 2.0

SeeDance 2.5 wins on duration and reference capacity: 30 seconds against 15, and up to 50 assets against 15. It also accepts audio-only reference requests and now reaches 1080p. SeeDance 2.0 Pro still wins on maximum output size because its Unifically route reaches 4K. SeeDance 2.0 also has the stronger public quality record today. On Arena's August 2 boards, its 720p build ranks #1 for image-to-video at 1478±10 and #2 for text-to-video at 1479±11.

When to use SeeDance 2.5

Use SeeDance 2.5 when a single video needs more than 15 seconds, a large reference set, or audio-only guidance. Use SeeDance 2.0 Pro when 4K output matters more than duration. For early drafts, compare the full job price on the pricing page: at 480p, 2.5 is cheaper per second with a reference video than without one.

What SeeDance 2.5 can do

Native dialogue and cinematic scene progression

A close-up radio detail expands into a storm-lit rescue scene, then lands on a spoken emotional beat. This tests coherent camera progression, weather, a human face, and synchronized dialogue in one short generation.

Product detail, glass, and liquid physics

A macro product orbit combines transparent glass, refracted light, moving water, and fine surface detail. It is a compact stress test for material realism and controlled commercial camera movement.

Fast action and camera handoff

The shot moves from wheel-level tracking to a whip-pan and aerial chase while the rider, bicycle, leaves, and landing stay in motion. This tests speed, articulated movement, continuity, and layered environmental sound.

FAQs

People also ask

Yes. The model is live on Unifically as bytedance/seedance-2.5, using the shared /v1/tasks generation flow.

Any whole number of seconds from 4 through 30. The default is 4 seconds.

The Unifically route supports 480p, 720p, and 1080p. ByteDance shows 4K output in its consumer product, but 4K is not callable on this API route.

Yes. Synchronized voice, sound effects, and background music are enabled by default. Set generate_audio to false for silent output.

Up to 30 images, 10 videos, and 10 audio files, with no more than 50 assets in one request. Video and audio references can each total up to 30 seconds.

Yes. Both frame URLs are required in frame mode, and the output follows the first frame's aspect ratio. Frame URLs cannot be mixed with reference arrays.

Yes. Use bytedance/seedance-2.5-draft (the Draft option in the playground) to generate a 480p draft. When you like it, send the draft's task ID as draft_task_id and the same video is generated natively in 1080p from that task, without an external upscaler. 720p and 4K renders from a draft are coming soon.

Pricing is per output second. Resolution and the presence of a reference video determine the rate; the pricing page shows the live matrix.

Guides and comparisons

Review pricing, limits, and tested outputs before running this model.