Docs

Music for video

This module adds music that is composed FOR your finished clip. The AI system watches your video, detects the tempo and the moments of change on its own, and composes three versions of a track exactly as long as the film. You play each version straight away on your own clip, with the picture, pick one, set the volume, and the bed drops under the voice whenever someone in the recording speaks. The finished file lands in your Asset Library and can go on to get a voiceover or captions.

What it's for

  • For whom: anyone preparing video for social media, a website or a presentation who wants a bed without digging through music libraries and without matching cuts by ear.
  • Use it when: you have a finished clip (your own, from the generator, or one that already has a voiceover) and you want the music to keep its rhythm.
  • If you want to add a voiceover or captions, use Dubbing + Lipsync - ideally ON the file with music, because the bed is then a separate track under the voiceover and ducks by itself.

Info

The music is written to the picture, not the other way round. The model sees your cuts and writes its tempo to them. It cannot, however, be handed a list of cuts - with a very dense edit it lands on the overall pulse and the strongest changes, not on every cut. That is why you get three versions to watch with the film, and the system suggests the one that sits best in the edit.

Before you start

  • Permissions: a role allowed to use services (member and up).
  • Module enabled: "Music for video" must be active for your company. If you do not see it under "Video editing", contact your WebImpact account manager.
  • Material: a video file from disk (MP4, MOV or WebM, up to 200 MB) or a clip from the Asset Library. The clip can be silent or have its own sound. One composition covers a clip up to the length set for your company (90 s by default).

How to add music

  1. In the company menu open Video editing → Music.
  2. Point at a clip - upload a file or pick one from the library (a film that already has a voiceover works too).
  3. Pick a music style from the list or leave "AI decides" - the model then describes the mood of the picture itself. Some styles have a short sample to listen to. If your company has custom descriptions enabled, describe the music you want in your own words, in any language - AI rewrites it in the background, while you type, into 3-5 English sound words (genre, bass, drums) and shows under the field exactly what the composer will get. Words about the song's arc (build-up, drop, intro) are dropped - they impose their own rhythm and break the fit to the cuts.
  4. Click Compose 3 versions. This uses one generation from the module's limit. Composition usually takes 1-3 minutes; you can refresh the page, nothing is lost.
  5. Watch the three versions. Each plays on your clip, with the picture and the music underneath - the way the finished film will look. We pre-select the one that sits best in the edit; keep it, or pick the one you simply like. The volume sliders from the mix drive the previews too.
  6. Set the mix: music volume, and for a clip with its own sound also that sound's volume plus the "duck music when someone speaks" switch. We nudge the track to the cuts ourselves, by a few dozen milliseconds, invisible to the eye.
  7. Click Render the film with music. The picture stays untouched (copied losslessly), only the sound changes. The file lands in the Asset Library, and from the result screen you can go straight to dubbing or captions.

Good to know

  • Three versions, because the model is not deterministic - the same clip and style give different music every time. Three tries are the only way to have a choice.
  • The preview is not the file yet. In the preview the music plays evenly; ducking under speech is heard only in the rendered film. If no version sits in the edit, compose again - that is three new versions.
  • The music is as long as the clip, so if you trim the film later in another module the track will be too long - compose it again then.
  • Order with a voiceover. Music first, then voiceover: the bed is then a separate track under the voice and ducks by itself. Voiceover first, then music also works - the bed then ducks under the voice in the recording.
  • What it costs. The limit counts generations (one composition = one slot) regardless of clip length. The mix itself does not use the limit - render as many times as you like.