Back to Blog
audio preparation for dubbingqualityhow-to

Audio Prep Checklist Before You Dub Anything

DubLab TeamAugust 28, 2026 13 min read

When your dubbing project begins, the quality of your source audio sets a ceiling on everything that follows. No amount of post-processing can recover a recording that was noisy to start, or clipped during capture. A clean take saves you time in translation and synthesis, and it shows in the final dub.

Before you hit record, run through this practical checklist. Most of these points take seconds to verify. They work whether you are recording new dialogue or extracting audio from an existing video.

Recording studio with microphone and acoustic panels

Why Audio Prep Matters Before Dubbing

The dubbing workflow has four stages: recording or extraction, processing, translation, and synthesis. Poor source audio creates problems at every stage after recording. When you record with clipping, that damage is permanent. When you record in a noisy room, the noise travels through every dubbed language. When your levels are inconsistent, synthesis tools have to guess the intended dynamics and often get it wrong.

Extraction from existing video introduces different challenges. If the original was compressed to MP3 or heavily processed, you are working with already-degraded audio. If the original video has music or background sound mixed with dialogue, you will need to separate them first. These extraction scenarios benefit even more from a checklist, because the source material is less under your control.

The difference between a 10-minute preparation and a 30-minute dubbing job with audio problems is enormous. Most audio issues are faster to prevent than to fix.

The Recording Setup Workflow

Before you touch your microphone, invest 5 minutes in setup. This workflow applies whether you are in a professional studio, an office, or a bedroom.

Step 1: Choose Your Room

The room itself is your first acoustic tool. Smaller rooms with soft surfaces absorb sound better than large, hard-walled spaces. A bedroom with curtains and a rug will sound cleaner than a conference room with tile and glass. A closet with clothes hanging around three walls is excellent if you do not mind the space.

Open the door, close the windows, and run a 10-second test recording at normal speaking volume. Listen on headphones. If you hear air conditioning, traffic, or a constant hum, make a note. Turn off what you can (AC, fans, nearby devices) and test again. If the noise drops, you found the source. If it persists, move to a different room and test there instead.

Step 2: Position the Microphone and Add a Pop Filter

Position your microphone 6 to 12 inches from the speaker's mouth. Closer than 6 inches risks plosives (hard p and b sounds) and proximity effect (a bass boost that muddies speech). Farther than 12 inches picks up more room noise and loses clarity. If you are using a headset mic, adjust the boom so it points at the corner of your mouth, not directly at your lips. This angle matters: plosives will disperse rather than slam the diaphragm.

A pop filter or windscreen is not optional if you are recording speech. These simple accessories cut down plosive bursts that otherwise clip the microphone. If you do not have a commercial pop filter, a foam ball or even a thin cloth held 2 inches in front of the mic works well enough. The goal is to diffuse the burst, not to block all sound.

Step 3: Set Levels and Do a Test Take

This step prevents the most common recording disaster: clipping. Set your recording level so the loudest parts of normal speech hit around 70 to 80 percent of your meter, not 100 percent. If you record too hot, your audio will clip, and clipping cannot be undone. It is permanent distortion that sounds like a harsh crackle or pop at the peak.

Do a test take at your normal speaking volume. Speak as you will during the real recording. Watch the levels. If peaks hit 90 percent or higher, lower your gain or move the microphone back slightly. Then do the real take. This 2-minute step prevents 30 minutes of regret.

Step 4: Record Consistently

Record your full script in one pass at consistent speaking pace and energy. Do not stop and start repeatedly. Stitching together fragments creates audible joins where tone, room sound, and breath pattern shift. One clean take, start to finish, is much easier to work with than five snippets spliced together.

If you stumble or a plane flies over, just pause for a beat and keep going. It is easier to edit out a gap than to splice mismatched recordings. Aim for two good full takes if you have time. You can pick the cleaner one or use the best segments from each.

Audio Failure Modes and How to Fix Them

Even careful recording can encounter problems. Knowing what to look for and how to respond saves time during the dub.

Consistent Hum or Buzz in the Background

This symptom is usually an electrical issue: a grounding loop or a nearby electrical appliance. You hear it as a low 50 Hz or 60 Hz tone, or as fluorescent buzz at higher frequencies. It is nearly impossible to remove without expensive audio tools and often sounds worse after processing.

Fix at the source: unplug devices near your recorder, move away from the power outlet, or use a different power circuit in your building. If the hum persists, try recording in a different room. If you catch this during playback but after recording, your best option is to re-record. Do not try to EQ it out during the dub; it will come back in unexpected ways across languages.

Clipping on Specific Words

Some words trigger louder spikes than others. Words with hard consonants at the start (power, pretty, back) or long vowels (go, home, we) sometimes peak higher than your test take did. When you hear distortion on specific words during playback, it is too late to EQ.

Prevention: watch your meters during the recording, not just during the test. If you see a spike exceed 85 percent, lower your gain slightly and do another take. If clipping did occur, re-record those lines at a lower level.

Inconsistent Energy or Volume Across the Recording

Your voice may drift lower or higher as you get tired or forget to maintain your distance from the mic. This is audible as a voice that starts clear and loud but trails off quiet by the end, or vice versa. Inconsistent energy confuses synthesis tools; they learn from the beginning of your recording and may mismatch the later sections.

Fix: watch your levels throughout the take and maintain consistent distance from the mic. If you notice drift during playback, mark the sections where it happens and re-record them. Do not try to normalize this in post unless you are confident with audio editing.

Muffled or Distant Sound

If your recording sounds like you are speaking from inside a pillow, the microphone is too far away, positioned incorrectly, or the room is too dead (over-damped with heavy curtains and acoustic foam). This is harder to fix in post than other issues and is better prevented by using proper mic placement and testing before the full take.

Fix: move the microphone closer (but not closer than 6 inches), position it at the correct angle, and do another test. If the room is the issue, open a window partially or move to a room with harder surfaces.

Excessive Breath or Mouth Sounds

Even careful recording picks up breath between phrases and slight mouth clicks. These are normal, but excessive breath or popping sounds suggest the microphone is too close or the pop filter is not working. This is audible and often surviving in the final dub if not addressed.

Fix: add or adjust your pop filter, move the microphone back 1 or 2 inches, and re-record. During the recording, breathe gently between phrases instead of deeply.

Worked Example: Recording a Product Demo Script

Let us walk through a real scenario. You are recording a 3-minute product demo script for YouTube. You want to dub it into Spanish, French, and German. Here is how the checklist applies.

Room selection: You have a small office and a bedroom. Test both. The office has open shelves and hard walls. The bedroom has curtains, a bed, and soft surfaces. The bedroom sounds quieter during your 10-second test, so you choose the bedroom. You close the door, draw the curtains, and move your desk chair to face the wall.

Microphone setup: You are using a USB microphone on a desk stand. You position it 9 inches from your mouth and adjust the pop filter to sit 2 inches in front. You do a test recording of the first sentence: "Welcome to the product demo." You listen and hear your voice clearly, no plosives on welcome, no background hum.

Level setting: You watch the meter. The loudest word (demo) peaks at 78 percent. That is in the target range. You do another test with the full first paragraph, speaking at a brisk, energetic pace like you plan to in the real take. Peaks stay between 70 and 82 percent. You are ready.

Recording: You record the full 3-minute script in one pass. You stumble slightly on a technical term at 1:45 but keep going. Pause for a beat, then continue. Your energy stays consistent.

Playback check: You listen to the entire recording on headphones. The stumble is there but short, editable. The voice is clear throughout. No hum, no clipping, no fading energy. You have one take you are happy with, but you record a second take anyway. This time you nail the technical term, but your energy feels slightly forced. You will use the first take.

Export: You export as WAV at 48 kHz, 24-bit. The file is 60 MB. You upload it to DubLab. The dubbing process handles the rest: translation to Spanish, French, and German, voice synthesis using your voice characteristics, and delivery in all three languages. Because your source audio was clean, the dubs sound natural and consistent.

This entire process, from room selection to export, took about 30 minutes. A sloppy recording at a low level with hum in the background would have required expensive cleanup or even re-recording later.

Format and Technical Settings

Export or record in WAV or AIFF, not MP3. Compressed formats throw away audio data that is hard to restore. Use 16-bit depth at a minimum, 24-bit is better. Sample rate should match your video: 48 kHz for broadcast and most platforms, 44.1 kHz for podcasts and music.

If your recorder has built-in compression or voice enhancement, turn it off. You want the raw signal so you have maximum flexibility in post-production. Many smartphone voice recorders enable voice enhancement by default; check settings and disable it.

Recording Setup Comparison

Different recording environments have different trade-offs. Here is how common setups compare:

SetupNoise LevelPlosive ControlCostTime to Set Up
Smartphone with built-in micHighPoor0 dollars2 minutes
Smartphone with budget USB micLowGood30-50 dollars5 minutes
Handheld digital recorderVery lowGood100-300 dollars5 minutes
USB mic in treated roomVery lowExcellent50-200 dollars10 minutes
Professional studio rentalExcellentExcellent50-150 dollars per hour0 minutes

For most people, a budget USB microphone in a quiet room will outperform a smartphone with the built-in mic, even if the room is not acoustically treated. The microphone matters more than the space when you are starting out.

Common Recording Mistakes and How to Avoid Them

Mistake 1: Recording Multiple Takes Back to Back Without Listening

If you record take 1, pause briefly, and record take 2 without listening to take 1, you might not notice that take 1 has clipping, hum, or a dropped word. You only catch the problem when you are ready to dub.

Fix: Always listen to your take immediately after recording, on headphones or a good speaker. A 30-second listen will catch 90 percent of problems before you move on.

Mistake 2: Using the Built-in Microphone on Your Computer or Phone

Laptop and phone microphones are optimized for calls, not for recording voice for dubbing. They have low sensitivity, add noise, and often have automatic gain control that changes levels mid-recording. You will always sound worse than with any external microphone.

Fix: Invest in even a basic USB microphone (under 50 dollars) or a handheld recorder. The sound quality improvement is dramatic.

Mistake 3: Not Accounting for Room Ambience in Extraction

If you are extracting audio from an existing video, the background room tone or silence might not match across cuts or different scenes. This is only visible if you look at the waveform over time.

Fix: When extracting, normalize the levels across different video segments. If combining audio from multiple sources, listen for consistency in background tone and try to align it. Some audio editors can do this automatically.

Mistake 4: Recording at Very Low Levels to Avoid Clipping

Some people record very quietly to be absolutely sure they will not clip. Then they amplify the audio in post. This amplifies noise along with the signal and adds hiss.

Fix: Record at 70 to 80 percent level. This is the right balance: you have headroom against clipping but you are not amplifying noise unnecessarily.

Pre-Dubbing Audio Checklist

Before you upload your audio for dubbing, run through this list on headphones or good speakers:

  • Voice is clear and natural, not strained, distant, or overly loud
  • No constant background hum, buzz, or air conditioning noise
  • No words sound distorted or clipped at the peaks
  • Energy and volume stay consistent from start to finish
  • No excessive breath, mouth clicks, or rustling between phrases
  • Format is WAV or AIFF at 16-bit or 24-bit, 48 kHz sample rate

If you notice any problem, re-record that section or the entire audio. A few minutes now saves hours of frustration and compromise later.

What to Do Next

  1. Locate or borrow a microphone. A basic USB mic, a smartphone with a recording app, or a handheld recorder will all work. Test it in your chosen room first.

  2. Run one test recording. Spend 10 minutes recording a short phrase (10 to 20 seconds), then listen critically. Notice hum, distance, breath, clipping, or any other issues. Adjust one thing at a time and test again.

  3. Record your full script. Use the workflow above: choose room, set levels, add pop filter, record consistently. Aim for two good takes.

  4. Listen to your recording on at least two different playback devices. Headphones and a speaker will reveal different problems. If something sounds wrong on both, it is a real issue worth re-recording.

  5. Export as WAV at 48 kHz, 16-bit or higher, then upload to DubLab. Your audio is now ready for translation and synthesis in 92+ languages.


🚀 Start Dubbing Your Videos Today

DubLab uses AI to translate your videos into 92+ languages in minutes.

📱 Download for iOS

🌐 Try Free at dublab.app