Models

Which models generate native synchronised audio

Model Seedance 2.5Model Veo 3.1Model Wan 2.7Format Short dramaMarket Global English
Short answer

All eighteen video models generate audio. Six give you a toggle to switch it off: Seedance 2.5, Seedance 2.0, Seedance 2 Fast, Seedance 2 Mini, Vidu Q3-Turbo and Kling 3.0 Omni. The other twelve have sound always on with no way to disable it. So the useful question isn't which models do native audio — it's which ones let you decline it.

Native audio is listed as a feature on almost every model in this category. The picker shows something more useful: which of them will let you turn it off.

Why does the toggle matter more than the capability?

Because you can't un-generate a soundtrack you didn't want.

If you've cast a voice deliberately — a licensed performer, a synthetic voice you chose, a recording you're syncing — then a model generating its own audio is producing a bed you'll be working against rather than with.

On a model with a toggle, that's one switch. On a model without one, you're mixing under audio you didn't ask for, and you can't remove it at the source.

Which inverts the usual framing. "Does this model do native audio" is a capability question with eighteen yeses. "Can I turn it off" is a workflow question with six.

Every model generates audio. Six let you decline it. That's the number that decides your workflow.

Which six have a toggle?

Four Seedance models plus two others.

ModelDurationResolutionAudio control
Seedance 2.54–30s480 / 720 / 1080Toggle
Seedance 2.04–15s480 / 720 / 1080 / 4KToggle
Seedance 2 Fast4–15s480 / 720Toggle
Seedance 2 Mini4–15s480 / 720Toggle
Vidu Q3-Turbo1–16s540 / 720 / 1080Toggle
Kling 3.0 Omni3–15s720 / 1080Toggle

Source: Hexcoded Creative Studio picker, Video → Create, September 2026.

Notice that the entire Seedance family has a toggle and no Google, Wan, MiniMax, Grok, LTX or Hailuo model does. That's a vendor pattern rather than a capability one.

And notice that Kling 3.0 Omni has a toggle while Kling 3.0 Turbo doesn't. Same family, different behaviour — Turbo's own description says "audio always on," which is the product telling you before you select it.

Which twelve are always on?

Everything else, and two of them are notable.

ModelAudioWhat its description says about sound
Gemini Omni FlashAlways on"Any input in, sound-matched video out"
Veo 3.1Always on"Native-audio film in 4K, extendable past a minute"
Veo 3.1 FastAlways on"Faster Veo, same sound and scene extension"
Veo 3.1 LiteAlways on"Veo quality, lighter and quicker"
MiniMax H3Always onNothing about sound
Wan 2.7Always onNothing about sound
Wan 2.6Always onNothing about sound
HappyHorse 1.1Always on"Native audio and lip-sync in one pass"
Grok Imagine 1.5Always on"Cinematic clips with synchronised audio"
LTX-2.3 FastAlways onNothing about sound
Hailuo 2.3Always onNothing about sound
Kling 3.0 TurboAlways on"Fast cinematic clips, audio always on"

Source: Hexcoded Creative Studio picker, Video → Create, September 2026.

Two things worth pulling out of that.

Only three of the twelve say so in the description. Kling 3.0 Turbo names it directly. Gemini Omni Flash, Veo 3.1, Veo 3.1 Fast, HappyHorse 1.1 and Grok Imagine 1.5 all mention sound as a feature without saying it can't be disabled. The other six say nothing about audio at all — you find out in the settings panel.

Gemini Omni Flash has no duration or resolution control either. Its whole panel is references, prompt, aspect ratio and sound — and the sound reads "always on." So it's the model with the fewest decisions to make, and audio isn't one of them.

Where this falls short. "Always on" is what the picker says. Whether the generated audio is usable, or whether it can be replaced downstream in your own edit, is a separate question this post doesn't answer — and the answer differs depending on whether a licensed performer appears.

The trade-off in detail

Can you just replace it afterwards?

Sometimes, and there's a hard constraint if a real person is on screen.

For a generated character, the audio is yours to handle however your edit needs.

Where a licensed human creator appears, Hexcoded's terms permit light edits only — trimming, cropping, captions, music, end-cards. Re-voicing a face outside the platform is prohibited, and it's listed as a prohibited use rather than merely a contract breach.

So "generate with audio always on, fix it later" isn't available for a licensed performer. If the voice matters and the performer is real, you need a model where audio can be switched off, or you need Talking Actors where the voice is a casting decision.

With a licensed performer, "fix the audio later" isn't a workflow. Re-voicing outside the platform isn't permitted.

How does the audio actually get generated?

For one model, Hexcoded's own documentation names the mechanism — which is more than any competitor comparison offers.

Seedance 2.5 uses an audio-visual architecture that processes visual frames and acoustic signals within the same generation pass, where traditional video generation creates silent footage first and overlays external audio afterwards. Visual movement and sound dynamics are calculated simultaneously, so footsteps, environmental ambience, surface impacts and spoken dialogue align with visible movement without a post-synchronisation step.

That's the difference between matched and synced. A matched take has nothing to reconcile; a synced one has something you maintain.

For the other seventeen models, the picker tells you whether audio is generated and whether you can disable it. It doesn't tell you how, and neither will we without a source.

How should you decide?

Three questions, in order.

1

Is the voice a casting decision?

A specific age, accent or register, or matching a performer. If yes, you need a model with a toggle — or Talking Actors, where the actor carries the voice.

2

Is a licensed human creator on screen?

Then re-voicing outside the platform isn't permitted, so the toggle isn't a convenience — it's the only route to controlling the audio at all.

3

Neither?

Native audio is the faster path. Mouth and sound produced together have nothing to reconcile, which removes a whole class of problem.

Audio also works as an input rather than only an output. Creative Studio's video engine accepts audio clips as references — used for rhythmic cues, acoustic tone and synchronisation context — so you can generate to a track rather than cutting to one afterwards.

All audio behaviours read from Hexcoded's Creative Studio picker in Video → Create mode in September 2026. Behaviour changes with model versions and differs by mode. Verify in the picker before relying on this.

The bottom line
  • All eighteen models generate audio. Only six let you switch it off
  • Those six: Seedance 2.5, Seedance 2.0, Seedance 2 Fast, Seedance 2 Mini, Vidu Q3-Turbo, Kling 3.0 Omni
  • The entire Seedance family has a toggle. No Google, Wan, MiniMax, Grok, LTX or Hailuo model does
  • Kling 3.0 Omni has a toggle; Kling 3.0 Turbo doesn't. Same family, different behaviour
  • Only three of the twelve always-on models say so in their description. The rest you find in the settings
  • With a licensed performer, re-voicing outside the platform isn't permitted. The toggle is the only control you have
  • Gemini Omni Flash has no duration or resolution control either. Audio is one of the few settings it has
  • Audio is an input too. Attach a track as a reference and generate to it rather than cutting to it

All eighteen video models in Hexcoded's Create mode generate audio. The more useful distinction is which let you disable it — six do: Seedance 2.5, Seedance 2.0, Seedance 2 Fast, Seedance 2 Mini, Vidu Q3-Turbo and Kling 3.0 Omni.

Those six. The other twelve have sound always on with no toggle in the settings panel. The whole Seedance family has a toggle, and no Google, Wan, MiniMax, Grok, LTX or Hailuo model does.

No. Omni has an audio toggle; Turbo has audio always on, and its own description says so — "fast cinematic clips, audio always on." Same family, different behaviour.

For a generated character, yes. Where a licensed human creator appears, no — Hexcoded's terms permit light edits only and re-voicing a face outside the platform is listed as a prohibited use. So if the voice matters and the performer is real, you need a model with a toggle.

For Seedance 2.5, Hexcoded's own documentation describes an audio-visual architecture processing visual frames and acoustic signals in the same generation pass, rather than creating silent footage and overlaying audio afterwards. For the other models, the picker tells you whether audio is generated and whether you can disable it, not how.

Yes. Creative Studio's video engine accepts audio clips as references, used for rhythmic cues, acoustic tone and synchronisation context — so you can generate to a track rather than editing to one afterwards.

Sound on, sound off, sound as input

Every model's audio behaviour shows in the picker before you generate — toggle or always on, per model. And audio references let you generate to a track rather than cut to one.

Open Creative Studio

More on model capability, access and rights in Models.