← Home
TOOL

By Valmera Editorial · Published · Updated

Add Voiceover to Video

Upload recorded narration and specify where it belongs in the current edit. The voiceover has its own level, and its ducking option can lower original program audio during its window. Set the music balance separately and check the complete preview.

It works on real footage and on canvas projects made from images — which is how you turn a folder of photos and a voice memo into a narrated video.

Add a Voiceover

Upload the video, upload the narration, describe the mix. Account creation and uploads are free; editing requires a subscription.

Create an account →
See pricing →

How to Add a Voiceover to a Video

  1. 1
    Upload your video and narration
    Upload a supported video within both the 14 GB and three-hour limits, then add your recorded narration as a supported audio file up to 50 MB per track. Accessible supported links can also supply media. Indexing and editing require a subscription and processing credits.
  2. 2
    Type the request
    Type something like "add my voiceover over the current edit, lower the original audio and keep music quiet beneath it", "mute the original audio and use the voiceover instead", or "make the voiceover louder and the music quieter". The agent places the track and mixes the layers.
  3. 3
    Preview, refine, export
    Check the preview, adjust with follow-ups — "start the narration at 0:03", "apply social mastering after balancing the layers" — then export the approved version as a final MP4 in Studio. The standard Valmera paid-plan policy omits overlay watermarks, but service settings can override it. Inspect the actual final file for branding and any end card.

Describe revisions in chat and check the resulting preview before exporting.

Balance narration, music and original audio

With voiceover ducking enabled, original program audio is lowered by 12 dB during the active voiceover window. This is separate from smooth music ducking, which uses program audio as its reference. Uploaded narration is mixed later and does not automatically drive that music sidechain. Listen to the music against the actual narration and adjust its level explicitly.

Each layer has a gain setting. Identify which layer needs changing and compare quiet speech, louder moments and gaps. Optional social mastering targets −14 LUFS with −2 dBTP headroom and output limiting after the mix; these are processing settings, not universal platform requirements. Check the delivered file. Full details in the audio and music docs.

Replace the Original Audio with Your Narration

Screen recordings, b-roll, gameplay, product demos — often the original sound just needs to go. Ask the agent to mute the original audio and it drops to full silence, leaving your voiceover as the voice of the video. Add a bed underneath — the agent finds a track on the web by genre or vibe — then review and set its balance under the narration.

Prefer to keep some of the original sound? Lower it instead of muting it — per-layer gain means "original audio at 20% under the voiceover" is a valid, one-sentence mix.

Narrated Slideshows: Voiceover Over Images

You don't need footage to need a voiceover. Valmera's canvas projects start with no video at all: images, clips, and music become a sequential timeline in any aspect ratio, stills get Ken Burns motion so they feel alive, and your narration plays over the top. Real-estate walkthroughs, photo essays, tutorials from screenshots — described in chat, assembled by the agent.

Honest scope: canvas projects don't get transcript-based captions, speed changes, or censor effects — manual captions, on-screen text, overlays, and music all work.

Bring Your Own Voice — Honest Limits

Three things this page won't pretend. Valmera does not generate AI voices or text-to-speech — the narration is yours, recorded by you. It does not offer denoise or "studio sound" processing — a noisy recording stays a noisy recording, so record as clean as you can. And it does not do per-speaker leveling within one track — layers are balanced against each other, not voices within a single file.

Audition the narration with the original audio and music. Confirm the selected recording, start point, level, and gaps, and check that speech remains intelligible. A recorded mix setting is useful evidence, but the preview and final MP4 are where you verify what the listener will hear.

More Than a Voiceover Tool

Once the narration is in, the same chat finishes the video: cut the dead air, style word-timed captions, add real recorded sound effects found on the web, reframe for vertical, grade the color. Browse all the AI video editing tools — one agent, every request.

Frequently Asked Questions

In Valmera you upload your video, add your recorded narration as an audio file (up to 50MB per track), and type what you want: "add my voiceover over the current edit, lower the original audio and keep music quiet beneath it". The agent places the narration on its own layer and mixes the other audio underneath. No timeline editing required — refinements are follow-up messages.
Do not assume it will. Voiceover ducking lowers original program audio during the voiceover window. Smooth music ducking uses program audio as its reference; the separate voiceover is mixed later. Set the music level against uploaded narration explicitly and listen to the relevant passages.
Yes. Ask the agent to mute the original audio — it goes fully to silence — and your voiceover becomes the voice of the video, with music found on the web or your uploads underneath if you want it. Each layer keeps independent volume control throughout.
Valmera does not offer text-to-speech narration. Supply recorded narration, then request placement, original-audio ducking and the required layer levels. Review the mix before applying any optional overall loudness treatment.
Yes — Valmera supports canvas projects that start with no video at all: images, clips, and music arranged as a sequential timeline, in any aspect ratio, with Ken Burns motion on stills. Add your voiceover on top and you have a narrated slideshow. Honest note: canvas projects don't get transcript-based captions or speed changes — manual captions, text, overlays, and music all work.
Studio accepts MP3, WAV, M4A, AAC and OGG audio up to 50 MB per track. Supported accessible links can also supply media, subject to availability and processing limits. Identify the imported recording before placing it and adjusting its volume.
Name the voiceover, music and original-audio layers separately and specify the intended balance. Listen to quiet and loud passages before accepting it. Optional social mastering targets −14 LUFS with −2 dBTP headroom and output limiting; it is not a universal platform requirement or a guarantee of the measured final output.

Add and review your narration

Place the recording, balance the audio layers and inspect the final mix. Account creation and uploads are free; editing requires a subscription.

Create an account →
See pricing →

Related Articles

Add Music to Video
Choose a track, set its timing and review its balance under speech.
Mute a Video
Silence the original audio or lower any layer independently.
Docs: Audio & Music
Every audio capability, parameter by parameter.
All AI Video Editing Tools
The full Valmera toolset, one chat away.