All eighteen video models generate audio. Six give you a toggle to switch it off: Seedance 2.5, Seedance 2.0, Seedance 2 Fast, Seedance 2 Mini, Vidu Q3-Turbo and Kling 3.0 Omni. The other twelve have sound always on with no way to disable it. So the useful question isn't which models do native audio — it's which ones let you decline it.
Native audio is listed as a feature on almost every model in this category. The picker shows something more useful: which of them will let you turn it off.
Why does the toggle matter more than the capability?
Because you can't un-generate a soundtrack you didn't want.
If you've cast a voice deliberately — a licensed performer, a synthetic voice you chose, a recording you're syncing — then a model generating its own audio is producing a bed you'll be working against rather than with.
On a model with a toggle, that's one switch. On a model without one, you're mixing under audio you didn't ask for, and you can't remove it at the source.
Which inverts the usual framing. "Does this model do native audio" is a capability question with eighteen yeses. "Can I turn it off" is a workflow question with six.
Every model generates audio. Six let you decline it. That's the number that decides your workflow.
Which six have a toggle?
Four Seedance models plus two others.
| Model | Duration | Resolution | Audio control |
|---|---|---|---|
| Seedance 2.5 | 4–30s | 480 / 720 / 1080 | Toggle |
| Seedance 2.0 | 4–15s | 480 / 720 / 1080 / 4K | Toggle |
| Seedance 2 Fast | 4–15s | 480 / 720 | Toggle |
| Seedance 2 Mini | 4–15s | 480 / 720 | Toggle |
| Vidu Q3-Turbo | 1–16s | 540 / 720 / 1080 | Toggle |
| Kling 3.0 Omni | 3–15s | 720 / 1080 | Toggle |
Source: Hexcoded Creative Studio picker, Video → Create, September 2026.
Notice that the entire Seedance family has a toggle and no Google, Wan, MiniMax, Grok, LTX or Hailuo model does. That's a vendor pattern rather than a capability one.
And notice that Kling 3.0 Omni has a toggle while Kling 3.0 Turbo doesn't. Same family, different behaviour — Turbo's own description says "audio always on," which is the product telling you before you select it.
Which twelve are always on?
Everything else, and two of them are notable.
| Model | Audio | What its description says about sound |
|---|---|---|
| Gemini Omni Flash | Always on | "Any input in, sound-matched video out" |
| Veo 3.1 | Always on | "Native-audio film in 4K, extendable past a minute" |
| Veo 3.1 Fast | Always on | "Faster Veo, same sound and scene extension" |
| Veo 3.1 Lite | Always on | "Veo quality, lighter and quicker" |
| MiniMax H3 | Always on | Nothing about sound |
| Wan 2.7 | Always on | Nothing about sound |
| Wan 2.6 | Always on | Nothing about sound |
| HappyHorse 1.1 | Always on | "Native audio and lip-sync in one pass" |
| Grok Imagine 1.5 | Always on | "Cinematic clips with synchronised audio" |
| LTX-2.3 Fast | Always on | Nothing about sound |
| Hailuo 2.3 | Always on | Nothing about sound |
| Kling 3.0 Turbo | Always on | "Fast cinematic clips, audio always on" |
Source: Hexcoded Creative Studio picker, Video → Create, September 2026.
Two things worth pulling out of that.
Only three of the twelve say so in the description. Kling 3.0 Turbo names it directly. Gemini Omni Flash, Veo 3.1, Veo 3.1 Fast, HappyHorse 1.1 and Grok Imagine 1.5 all mention sound as a feature without saying it can't be disabled. The other six say nothing about audio at all — you find out in the settings panel.
Gemini Omni Flash has no duration or resolution control either. Its whole panel is references, prompt, aspect ratio and sound — and the sound reads "always on." So it's the model with the fewest decisions to make, and audio isn't one of them.
Where this falls short. "Always on" is what the picker says. Whether the generated audio is usable, or whether it can be replaced downstream in your own edit, is a separate question this post doesn't answer — and the answer differs depending on whether a licensed performer appears.
Can you just replace it afterwards?
Sometimes, and there's a hard constraint if a real person is on screen.
For a generated character, the audio is yours to handle however your edit needs.
Where a licensed human creator appears, Hexcoded's terms permit light edits only — trimming, cropping, captions, music, end-cards. Re-voicing a face outside the platform is prohibited, and it's listed as a prohibited use rather than merely a contract breach.
So "generate with audio always on, fix it later" isn't available for a licensed performer. If the voice matters and the performer is real, you need a model where audio can be switched off, or you need Talking Actors where the voice is a casting decision.
With a licensed performer, "fix the audio later" isn't a workflow. Re-voicing outside the platform isn't permitted.
How does the audio actually get generated?
For one model, Hexcoded's own documentation names the mechanism — which is more than any competitor comparison offers.
Seedance 2.5 uses an audio-visual architecture that processes visual frames and acoustic signals within the same generation pass, where traditional video generation creates silent footage first and overlays external audio afterwards. Visual movement and sound dynamics are calculated simultaneously, so footsteps, environmental ambience, surface impacts and spoken dialogue align with visible movement without a post-synchronisation step.
That's the difference between matched and synced. A matched take has nothing to reconcile; a synced one has something you maintain.
For the other seventeen models, the picker tells you whether audio is generated and whether you can disable it. It doesn't tell you how, and neither will we without a source.
How should you decide?
Three questions, in order.
Is the voice a casting decision?
A specific age, accent or register, or matching a performer. If yes, you need a model with a toggle — or Talking Actors, where the actor carries the voice.
Is a licensed human creator on screen?
Then re-voicing outside the platform isn't permitted, so the toggle isn't a convenience — it's the only route to controlling the audio at all.
Neither?
Native audio is the faster path. Mouth and sound produced together have nothing to reconcile, which removes a whole class of problem.
All audio behaviours read from Hexcoded's Creative Studio picker in Video → Create mode in September 2026. Behaviour changes with model versions and differs by mode. Verify in the picker before relying on this.
- All eighteen models generate audio. Only six let you switch it off
- Those six: Seedance 2.5, Seedance 2.0, Seedance 2 Fast, Seedance 2 Mini, Vidu Q3-Turbo, Kling 3.0 Omni
- The entire Seedance family has a toggle. No Google, Wan, MiniMax, Grok, LTX or Hailuo model does
- Kling 3.0 Omni has a toggle; Kling 3.0 Turbo doesn't. Same family, different behaviour
- Only three of the twelve always-on models say so in their description. The rest you find in the settings
- With a licensed performer, re-voicing outside the platform isn't permitted. The toggle is the only control you have
- Gemini Omni Flash has no duration or resolution control either. Audio is one of the few settings it has
- Audio is an input too. Attach a track as a reference and generate to it rather than cutting to it
All eighteen video models in Hexcoded's Create mode generate audio. The more useful distinction is which let you disable it — six do: Seedance 2.5, Seedance 2.0, Seedance 2 Fast, Seedance 2 Mini, Vidu Q3-Turbo and Kling 3.0 Omni.
Those six. The other twelve have sound always on with no toggle in the settings panel. The whole Seedance family has a toggle, and no Google, Wan, MiniMax, Grok, LTX or Hailuo model does.
No. Omni has an audio toggle; Turbo has audio always on, and its own description says so — "fast cinematic clips, audio always on." Same family, different behaviour.
For a generated character, yes. Where a licensed human creator appears, no — Hexcoded's terms permit light edits only and re-voicing a face outside the platform is listed as a prohibited use. So if the voice matters and the performer is real, you need a model with a toggle.
For Seedance 2.5, Hexcoded's own documentation describes an audio-visual architecture processing visual frames and acoustic signals in the same generation pass, rather than creating silent footage and overlaying audio afterwards. For the other models, the picker tells you whether audio is generated and whether you can disable it, not how.
Yes. Creative Studio's video engine accepts audio clips as references, used for rhythmic cues, acoustic tone and synchronisation context — so you can generate to a track rather than editing to one afterwards.
Sound on, sound off, sound as input
Every model's audio behaviour shows in the picker before you generate — toggle or always on, per model. And audio references let you generate to a track rather than cut to one.
Open Creative StudioMore on model capability, access and rights in Models.