All posts
Model Launches

Seedance 2.5 on Flixly: 30-Second Shots with Synchronized Audio

Seedance 2.5 is ByteDance's newest video model — single-shot clips up to 30 seconds with audio generated in the same pass, from a prompt, a start frame, or up to 50 references. Here's how to run it on Flixly, and what it costs.

August 7, 2026
Seedance 2.5 on Flixly: 30-Second Shots with Synchronized Audio

TL;DR

Seedance 2.5 is ByteDance's newest video model. It generates single-shot clips of 4 to 30 seconds at 480p or 720p with synchronized audio rendered in the same pass, and accepts a text prompt, a start frame (with an optional end frame), or up to 50 image, video, and audio references. All three modes are live on Flixly today.

Seedance 2.5 is ByteDance's newest video model, and it is live on Flixly in all three modes: text-to-video, image-to-video, and reference-to-video.

What changed since Seedance 2.0

  • Up to 30 seconds in one generation. Seedance 2.0 topped out at 15. Thirty seconds is long enough for a complete beat — a setup, a turn, and a landing — without stitching separate clips and fighting continuity between them.
  • Audio in the same pass. Dialogue, music, ambience, and effects are generated with the picture rather than dubbed over it afterward, so timing and lip movement line up. Put spoken lines in double quotes in your prompt.
  • Up to 50 references. Reference-to-video takes images, video clips, and audio together — up to fifty items — and you address them directly in the prompt as [Image1], [Video1], [Audio1].

Specs

Property Value
Modes on Flixly Text-to-Video, Image-to-Video, Reference-to-Video
Duration 4–30 seconds
Resolution 480p or 720p
Aspect ratios auto, 21:9, 16:9, 4:3, 1:1, 3:4, 9:16
Audio generated in the same pass, no extra cost
References up to 50 images, videos, and audio combined
Prompt length up to 5,000 characters

How to use it

  1. Open the Video Generator in the dashboard and pick the tab you need — Text, Image, or Reference.
  2. Pick Seedance 2.5 from the model selector.
  3. Set resolution and duration. Draft at 480p and short lengths — it costs less than half of 720p — then re-run the prompt you like at 720p and full length.
  4. Write the prompt as direction, not description: the subject, the action, the camera move, and any dialogue in double quotes.
  5. For image-to-video, upload a start frame; add an end frame if you want the shot to land on a specific composition.

Two things worth knowing about cost

Seedance 2.5 is billed by output length, so a 30-second clip costs roughly six times a 5-second one. The credit estimate on the page always reflects the exact settings you've chosen before you generate.

The second is less obvious: in reference-to-video, a video reference adds its own length to what you're billed for. A 20-second reference driving a 6-second output is billed on 26 seconds, not 6. Image and audio references don't do this — only video.

Writing a prompt for it

Seedance 2.5 responds to direction, not description. A prompt that reads like a shot list beats one that reads like a caption.

  • Say what moves. "A chef plates a dish" gives the model nothing to animate. "A chef sets the last herb on the plate, then slides it across the pass as the camera pushes in" gives it a beginning, an action and an end.
  • Put dialogue in double quotes. Anything quoted is generated as spoken audio, lip-synced in the same pass. She turns to the window and says "we should have left an hour ago" gets you the line and the mouth movement together.
  • Direct the sound. Ambience, score and effects are all generated. Naming them — rain on glass, a low sustained string, the click of a latch — is the difference between a scene with sound and a scene with room tone.
  • Address references by index. In reference-to-video the prompt refers to your uploads as [Image1], [Video1], [Audio1]. Say what each one is for: "the woman from [Image1], moving with the camera rhythm of [Video1]."

Which mode to pick

You have Use What it gives you
Only an idea Text to Video The model chooses framing and staging from the prompt alone
A starting frame Image to Video Your image becomes frame one — add an end frame to control where the shot lands
A character, product or style to hold Reference to Video Up to 50 references keep identity, motion or voice consistent across shots

If you're deciding between Seedance 2.5 and Seedance 2.0, the trade is length and references against resolution: 2.5 goes to 30 seconds and 50 references but stops at 720p, while 2.0 reaches 1080p and 4K in a 15-second window.

What it costs, and the one surprise

Seedance 2.5 is billed by output length, so a 30-second clip costs roughly six times a 5-second one, and 480p runs a little under half of 720p. The cheapest way to work is to draft short at 480p, then re-run only the prompt you like at full length and 720p.

The surprise is in reference-to-video: a video reference adds its own length to the bill. A 20-second clip driving a 6-second output is billed on 26 seconds, not 6. Image and audio references don't do this. Trim reference clips to the part that actually carries the motion you want.

Going deeper

For the full spec sheet — what ByteDance confirmed, what the API settled, and which claims circulating about this model are still unsupported — read Seedance 2.5: What's Real, What Isn't, and How to Run It.

Frequently Asked Questions

What is Seedance 2.5?

Seedance 2.5 is ByteDance's latest video generation model. Its headline change over Seedance 2.0 is length: a single generation runs up to 30 seconds instead of 15, and audio is generated together with the video rather than added afterward.

How long can a Seedance 2.5 video be?

Anywhere from 4 to 30 seconds, set in one-second steps. Longer clips cost proportionally more because the model is billed by output length, so start short while you iterate on the prompt and extend once the shot is right.

What resolutions and aspect ratios does it support?

480p or 720p, in 21:9, 16:9, 4:3, 1:1, 3:4, or 9:16 — or leave the aspect ratio on auto and the model matches your input. 480p costs less than half of 720p, which makes it the sensible tier for drafts.

Does Seedance 2.5 generate sound?

Yes. Dialogue, music, ambience, and effects come out of the same pass as the picture, so lip movement and timing line up without a separate voice step. Put spoken lines in double quotes in your prompt. Audio adds nothing to the cost.

What can I use as a reference?

In Reference to Video you can supply images, video clips, and audio — up to 50 references in total — and address them in the prompt as [Image1], [Video1], [Audio1]. Images anchor a character or product, video carries motion or style, audio drives rhythm and voice. Note that supplying a video reference adds that clip's length to what you're billed for.

Where can I try it on Flixly?

Seedance 2.5 is live in Text to Video, Image to Video, and Reference to Video on the Flixly dashboard. Pick it from the model selector, set your resolution and duration, and generate.

Tools mentioned in this post

SeedanceSeedance 2.5ByteDanceText to VideoImage to VideoReference to VideoAI VideoFlixly

Ready to create with Model Launches?

Jump straight into Flixly's AI studio and try model launches with 50+ models — free to start.