Seedance 2.5 Audio: Sound Generated in the Same Pass, at No Surcharge
Most AI video with sound is really two jobs: a silent render, then a separate voice or music pass laid on top. Seedance 2.5 does it in one. The generation request carries a generate_audio flag that defaults to true, and the audio comes back attached to the clip from the same job that produced the picture. VdoBloom sends that flag as true on every Seedance 2.5 render in every tab, so a 2.5 clip arrives with sound unless you strip it yourself in an editor.
The part worth knowing for budgeting is that sound is free of charge here. Seedance 2.5 bills strictly per second — 9.2 credits per second at 480p, 20.0 at 720p — and that rate is the whole price. A 5-second clip is 46 or 100 credits, a 30-second clip is 276 or 600, and none of those numbers move because audio is included. That is not universal on the platform: Seedance 1.5 Pro charges for its audio toggle, where a 5-second 720p clip goes from 8 credits silent to 15 with sound.
What you can steer, and how, depends on the tab. In Text to Video and Image to Video the prompt is yours — up to 30,000 characters — and the prompt is the only lever on the audio, because there is no dialogue field, no voice picker and no music selector anywhere in the flow. Describe the room, the delivery and the sounds you want and the model interprets that alongside the picture. On the Seedance Effects tab the effect's own prompt does that describing instead, so there you choose the sound by choosing the effect and the source photo.
What this tool does
- generate_audio defaults to true on Seedance 2.5, and VdoBloom sends it as true on every render in every tab — there is no audio switch to forget and none to find.
- Audio carries no surcharge on Seedance 2.5. The rate is 9.2 credits per second at 480p and 20.0 at 720p with sound or without: 5 seconds is 46 or 100 credits, 10 seconds 92 or 200, 30 seconds 276 or 600.
- Sound is produced by the video model in the same generation as the picture — one job, not a dubbing pass bolted on afterwards.
- Because Seedance 2.5 can render a single 30-second take, its audio arrives as one continuous 30-second pass rather than six clip-length takes that have to be crossfaded together.
- The prompt is the only audio control. In Text to Video and Image to Video you write it, up to 30,000 characters; on the Seedance Effects tab the effect supplies it. There is no dialogue field, no voice selector and no music picker in any of the three.
- Charging for audio is the norm elsewhere in the Seedance family's older tiers: Seedance 1.5 Pro costs 8 credits for a silent 5-second 720p clip and 15 credits for the same clip with audio.
- Other native-audio families on VdoBloom include the rest of the Seedance 2 line (Seedance 2, 2 Fast and 2 Mini), Wan 2.7 and Wan 2.6 Flash, MiniMax H3, SkyReels V4, PixVerse, Vidu Q3 and Q3 Turbo, and Lightricks LTX-2 Fast.
- Reference audio is supported for Seedance 2.5 in the main composer: the Seedance settings panel in Text to Video and Image to Video takes reference audio URLs. It cannot be combined with a first-frame or last-frame image in the same job, and the Seedance Effects flow does not send the field at all.
- The cheapest way to hear what Seedance 2.5 audio sounds like is a 4-second 480p render at 37 credits, which is more than the 10 free credits a new account starts with.
How it works
1.Open Text to Video if you want to describe the sound yourself
Go to /dashboard/video-creation/text-to-video/ and pick Seedance 2.5 from the Seedance group. This is the flow where the prompt is yours, which matters for audio because the prompt is the only thing steering it. Image to Video works the same way with a photo attached; the Seedance Effects tab replaces your prompt with the effect's.
2.Write the sound into the prompt, not just the picture
Say what the space sounds like, what the subject is doing with their voice, and what the ambience is under it. A prompt that only describes framing gives the audio pass nothing to work from. Actions with an obvious sound — footsteps, handling an object, speaking to camera — give the model more to score than a static portrait does.
3.Or choose an effect for the sound it implies
On the Seedance Effects tab the effect's stored prompt is what the model reads, so the effect is your audio direction. Pick on the strength of the action rather than only the visual style, and read what each live effect describes. Each one also states how many photos it needs from you, and the request is rejected with a specific count if you send the wrong number.
4.Set duration and resolution — audio scales with them at no extra cost
Every whole second from 4 to 30 is available, at 480p or 720p. The audio length simply follows the video length, and neither resolution changes what audio costs, because it costs nothing. That makes 480p the sensible place to audition sound: 10 seconds at 480p is 92 credits against 200 at 720p for the same audio pass.
5.Audition short, then commit
Run 4 to 6 seconds at 480p first, for 37 to 55 credits, and listen before you buy length. If the sound is wrong the fix is a rewritten prompt, a different effect or a different source photo — not a longer render. Credits are checked against your balance before submission and refunded automatically if the job fails.
Frequently asked questions
Does Seedance 2.5 generate audio?
Yes, natively and in the same generation as the video. The generate_audio parameter defaults to true and VdoBloom sends it as true on every Seedance 2.5 render, so clips come back with sound attached rather than needing a separate voice or music job afterwards.
Does Seedance 2.5 charge extra for audio?
No. Pricing is linear per second regardless — 9.2 credits per second at 480p and 20.0 at 720p — so a 5-second clip is 46 or 100 credits and a 30-second clip is 276 or 600 whether audio is generated or not. There is no separate audio line item and no rate multiplier, which is not true everywhere: Seedance 1.5 Pro's audio toggle takes a 5-second 720p clip from 8 credits to 15.
Can I turn the audio off?
Not from the interface. The underlying API accepts generate_audio as a parameter, but VdoBloom's Seedance flows send it as true on every render, so in practice Seedance 2.5 clips generated here always come with sound. Since audio adds nothing to the cost, the practical answer if you want a silent clip is to mute or drop the track in your editor.
Can I make the subject say a specific line?
You can ask for it in the prompt, but you cannot script it in a dedicated field. In Text to Video and Image to Video the prompt is yours and Seedance 2.5 accepts up to 30,000 characters, so a line of dialogue can be written into the brief — the model then decides the voice and the delivery, because there is no voice picker and no separate dialogue input. On the Seedance Effects tab you have even less say: the effect's own prompt is what runs.
Which other models on VdoBloom generate audio?
Native audio is available across several families, including the rest of the Seedance 2 line (Seedance 2, 2 Fast and 2 Mini), Wan 2.7 and Wan 2.6 Flash, MiniMax H3, SkyReels V4, PixVerse, Vidu Q3 and Q3 Turbo, and Lightricks LTX-2 Fast. Seedance 1.5 Pro has audio as a priced toggle rather than an included feature. What distinguishes Seedance 2.5 is not that it has sound but that it can carry 30 continuous seconds of it.
Is a 30-second clip one continuous audio track?
Yes. Because Seedance 2.5 renders 30 seconds as a single generation, the audio is one pass across the whole clip. The alternative route to 30 seconds — six 5-second renders — produces six independent audio takes at exactly the same credit cost (276 at 480p, 600 at 720p), plus the crossfade work to make them sound continuous.
Can I supply my own reference audio?
Yes, in the main composer. The Seedance settings panel in Text to Video and Image to Video has a reference audio URLs field that accepts public http(s) links, and the service forwards them for Seedance 2.5. One restriction applies: reference audio belongs to the multimodal input mode, so it cannot be combined with a first-frame or last-frame image in the same job. The Seedance Effects flow does not send the field at all, so effect renders generate their audio entirely from the effect's prompt and your images.
Related tools
Seedance 2.5: the 30-second model, and how to run it
ByteDance Seedance 2.5 renders 4-30 seconds at 480p or 720p only. Pick it in Text to Video, Image to Video or Effects. 30s at 480p is 276 credits.
Seedance 2.5 price: the full credit table, per second and per clip
Seedance 2.5 costs 9.2 credits/sec at 480p and 20 at 720p. Cheapest render is 37 credits (4s, 480p); a full 30s at 720p is 600 credits.
Seedance 2.5 Reference Images: How @Image1 and @Image2 Get Assigned
For Seedance 2.5, VdoBloom prepends "Reference @Image1 for the person's face" when an effect prompt has no @Image tag. Your uploads are numbered first.
30-Second AI Video Generator: The Longest Single Clip on VdoBloom
ByteDance Seedance 2.5 renders a full 30 seconds in one generation: 276 credits at 480p, 600 at 720p. The next longest on VdoBloom is 16 seconds.
Long AI Video Generator: One 30-Second Take Instead of Six Stitched Clips
One 30-second Seedance 2.5 take costs 276 credits at 480p — exactly what six stitched 5-second renders cost, without the drift between them.
Seedance 2.5 vs Seedance 2: duration against resolution
Seedance 2.5 runs 30s but caps at 720p; Seedance 2 stops at 15s and reaches 4K. 15s at 720p costs 300 vs 244 credits. The full maths.
Ready to try it?
Open Text to Video