Add Voiceover to Video
Upload recorded narration and specify where it belongs in the current edit. The voiceover has its own level, and its ducking option can lower original program audio during its window. Set the music balance separately and check the complete preview.
It works on real footage and on canvas projects made from images — which is how you turn a folder of photos and a voice memo into a narrated video.
Add a Voiceover
Upload the video, upload the narration, describe the mix. Account creation and uploads are free; editing requires a subscription.
Create an account →How to Add a Voiceover to a Video
- 1Upload your video and narrationUpload a supported video within both the 14 GB and three-hour limits, then add your recorded narration as a supported audio file up to 50 MB per track. Accessible supported links can also supply media. Indexing and editing require a subscription and processing credits.
- 2Type the requestType something like "add my voiceover over the current edit, lower the original audio and keep music quiet beneath it", "mute the original audio and use the voiceover instead", or "make the voiceover louder and the music quieter". The agent places the track and mixes the layers.
- 3Preview, refine, exportCheck the preview, adjust with follow-ups — "start the narration at 0:03", "apply social mastering after balancing the layers" — then export the approved version as a final MP4 in Studio. The standard Valmera paid-plan policy omits overlay watermarks, but service settings can override it. Inspect the actual final file for branding and any end card.
Describe revisions in chat and check the resulting preview before exporting.
Balance narration, music and original audio
With voiceover ducking enabled, original program audio is lowered by 12 dB during the active voiceover window. This is separate from smooth music ducking, which uses program audio as its reference. Uploaded narration is mixed later and does not automatically drive that music sidechain. Listen to the music against the actual narration and adjust its level explicitly.
Each layer has a gain setting. Identify which layer needs changing and compare quiet speech, louder moments and gaps. Optional social mastering targets −14 LUFS with −2 dBTP headroom and output limiting after the mix; these are processing settings, not universal platform requirements. Check the delivered file. Full details in the audio and music docs.
Replace the Original Audio with Your Narration
Screen recordings, b-roll, gameplay, product demos — often the original sound just needs to go. Ask the agent to mute the original audio and it drops to full silence, leaving your voiceover as the voice of the video. Add a bed underneath — the agent finds a track on the web by genre or vibe — then review and set its balance under the narration.
Prefer to keep some of the original sound? Lower it instead of muting it — per-layer gain means "original audio at 20% under the voiceover" is a valid, one-sentence mix.
Narrated Slideshows: Voiceover Over Images
You don't need footage to need a voiceover. Valmera's canvas projects start with no video at all: images, clips, and music become a sequential timeline in any aspect ratio, stills get Ken Burns motion so they feel alive, and your narration plays over the top. Real-estate walkthroughs, photo essays, tutorials from screenshots — described in chat, assembled by the agent.
Honest scope: canvas projects don't get transcript-based captions, speed changes, or censor effects — manual captions, on-screen text, overlays, and music all work.
Bring Your Own Voice — Honest Limits
Three things this page won't pretend. Valmera does not generate AI voices or text-to-speech — the narration is yours, recorded by you. It does not offer denoise or "studio sound" processing — a noisy recording stays a noisy recording, so record as clean as you can. And it does not do per-speaker leveling within one track — layers are balanced against each other, not voices within a single file.
Audition the narration with the original audio and music. Confirm the selected recording, start point, level, and gaps, and check that speech remains intelligible. A recorded mix setting is useful evidence, but the preview and final MP4 are where you verify what the listener will hear.
More Than a Voiceover Tool
Once the narration is in, the same chat finishes the video: cut the dead air, style word-timed captions, add real recorded sound effects found on the web, reframe for vertical, grade the color. Browse all the AI video editing tools — one agent, every request.
Frequently Asked Questions
Add and review your narration
Place the recording, balance the audio layers and inspect the final mix. Account creation and uploads are free; editing requires a subscription.
Create an account →