Skip to content

Wan 2.2 / Text to video

Add sound to a Wan 2.2 video without regenerating it

A silent Wan clip can still be a useful production asset. Build its soundtrack in an editor using audio you have the rights to use. This separates visual approval from sound design and makes revisions easier when the picture is already strong.

Break the soundtrack into layers

Begin with a quiet background bed that establishes the place. Add a few effects tied to visible actions, then add music only if it supports the message. A boat scene might need water and motor sound; it does not need a dramatic impact on every cut.

Watch the clip silently and mark the moments that need sound. If there is no clear visible event, a restrained ambience may be sufficient. Avoid adding speech that makes a fictional person appear to give a real endorsement.

Build the mix

  1. Import the approved Wan video and confirm whether the file actually contains audio.
  2. Add licensed or original ambience beneath the full shot.
  3. Place specific effects near their visual causes and adjust their level.
  4. Add fades, export and listen to the final file on headphones and a phone speaker.

Example sound plan

For a boat turning on open water, use a continuous low motor layer, a stronger wash of spray during the turn and a soft water bed across the cut. Keep narration intelligible if the clip appears in a narrated video. Do not treat this plan as audio that Wan automatically generated.

Preserve flexibility

Keep separate audio tracks in the project and save the original silent file. If the client changes the music, you should not need another video generation. When publishing a visual demonstration, describe whether the soundtrack is model-generated or added in post so viewers understand what the tool itself produced.

REVIEWED JOLLYAI IMPLEMENTATION

The workflow behind the guide

Written scene → configured Wan text conditioning and sampling → frame decoding → video encoding and any selected post-processing.

Inputs, presets & access
PRO for the reviewed standard Wan T2V route. Use the visible preset; do not infer file metadata from the prompt.
Implementation note
JollyAI runs configured Wan workflows with locally selected checkpoints and adapters. T2V, I2V and first/last-frame are separate routes. Presets and post-processing differ by route. Plan sound separately for the demonstrated silent workflow, and inspect the downloaded file before quoting dimensions or duration.

Configuration reviewed September 20, 2026. Your current playground controls and plans page determine availability. Prompt recipes above are suggestions, not guarantees or claimed test results.

Open Text to video ↗

Sources and useful next steps

The implementation notes describe JollyAI's reviewed local integration. Upstream capabilities and licences may differ from the settings available here.

Found an outdated setting or a reproducible problem? Contact JollyAI with the workflow and job reference. Keep account tokens and private input files out of public reports.