Video summary
Dialogue EQ Secrets for Film & YouTube
Main summary
Key takeaways
Main Ideas, Concepts, and Lessons (EQ Dialogue)
1) Why EQ Dialogue Is Necessary (Creator + Mixer Perspective)
- Raw dialogue recordings usually need EQ to sound full, rich, clear, and professional.
- EQing dialogue is harder than it seems because you must be able to hear what’s wrong—specifically, identify frequency balance deficiencies.
- EQ should be about moving from point A to point B, not applying presets or doing “random EQ.”
- Editing affects EQ workload:
- A good dialogue edit (tight cuts, consistent continuity) reduces problems caused by microphone placement, mic choice, and room tone changes.
- A poor edit forces you into more drastic EQ/noise reduction later.
- EQ can be done with any tool (stock or paid): the method is what matters.
2) Approach: Work Frequency Bands From Low → Mid → High
The speaker frames corrections by frequency region:
- Lows: rumble/mud
- Mids: boxiness, phasing, harshness around intelligibility
- Highs: ice-pick harshness + intelligibility without sounding brittle
Core principle: use your ears—because every mic/room/actor differs.
Detailed Methodology / Step-by-Step
A) Low-Frequency Problem Solving (Boomy/Muddy Dialogue)
Common Symptoms
- “Too boomy” / “too muddy”
- Often caused by overemphasized low frequencies, not by missing highs.
Core Fixes
-
High-pass filter (HPF / low-cut) first
- Purpose: remove rumble/hum/sub-bass that doesn’t support dialogue clarity.
- Avoid “eating into” the speaker’s fundamentals:
- Example: if a fundamental is around ~70 Hz, raising the HPF to ~100 Hz with a 12 dB/octave slope can reduce mud while preserving tone.
- Prefer gentle slopes (6 or 12 dB per octave) instead of brick-wall cuts.
-
Cut setting guidance
- For many creators/podcasters, you may need to cut higher than you expect since viewers don’t need sub-bass in spoken dialogue.
- Example guidance: could be as high as ~200 Hz depending on taste/content.
Why It Matters (Beyond Tone)
- Viewer devices: phones/tablets/TV make sub-bass unnecessary and potentially problematic.
- Headroom/bit-depth analogy: carving lows creates usable space so mids/highs get more of the “detail budget” in the digital file.
- Plosives: HPF can reduce some “P” pops close to the mic (proximity effect + no pop filter).
Optional Pro-Level Enhancement: HPF + Low Shelf
- Combine:
- HPF to remove mud
- Low shelf to restore upper bass body (often ~100–200 Hz)
- Use caution: too much boost can cause resonances.
- Goal: sound less thin even after reducing lows.
If You Want Perceived Low End Back (Thin Clips)
Tools mentioned:
- Low-target saturation (iZotope workflow noted; in Pro Tools saturation may be more full-band unless specialized tools are used)
- Subharmonic generation (warning: can sound artificial like LFE if overdone)
- Example: Waves R-Bass for perceived low end without full rumble.
Pro Tools Practical Suggestion
- Starting chain example:
- 6 dB/octave HPF
- Optionally a low shelf near 0 initially, then adjust
- Use automation for scene/segment changes
B) Midrange Cleanup (Trickiest Area—Especially for Lavaliers)
Typical Issues
- Mids are described as the most difficult band.
- Lav mic placement is often the primary offender:
- Capsules often sit near the zipper/chest (jaw/chin “shadow” effects).
- This can cause phasiness or strange midrange behavior.
Where to Look
Common problem areas (not “always”):
- Around ~200–500 Hz
- Or sometimes ~180–250 Hz or ~400–600 Hz
The guidance is to find what’s offensive for that mic/actor/costume.
Correction Technique
- Use small notch dips only where needed.
- Workflow trick:
- Loop a problematic phrase, raise gain to find the resonant frequency, then dip by a small amount (often 1–2 dB).
- Avoid over-EQing:
- Don’t create an EQ curve with lots of boosts/cuts.
- Over-correction can create artifacts (the speaker references “hair” / comb-filter-like issues and asks viewers to identify a specific artifact).
When Combining Mics (Shotgun + Lav)
- If shotgun + lav are both used:
- The lav may require less mid EQ because it supports the shotgun/broadcast sound.
- If it’s lav-only:
- Expect more cleanup.
Don’t Change “Identity”
- Don’t remove a speaker’s character just to match other voices—clarity should be balanced with natural timbre.
Optional: Boosting Mids (Less Common)
- Sometimes appropriate if a mic lacks mids and you need dialogue clarity/power.
- Risks:
- Avoid making lavs honky
- Avoid amplifying unwanted resonances
- Mentioned: some large diaphragm condensers may benefit from slight mid boost because cheaper LDCs can sound “smiley” (too much high/low relative to mids).
Advanced/Optional Tools
-
Dynamic EQ / multiband dynamics
- Use sparingly for resonances that poke out on specific words.
- Approach referenced in FabFilter: right-click to make dynamic and choose comp/expansion behavior.
-
EQ matching
- Harder in mids due to multiple sub-regions (woodiness vs upper mids).
- Mentioned example: “upper mid bite” around ~1–2 kHz; sometimes a small cut (e.g., ~3 dB) removes bite (example: nasal female voices).
-
Mid/side EQ
- Dialogue is mostly treated as mono, so mid/side EQ is generally less applicable, but discussion is welcomed.
C) High-Frequency Handling (Clarity Without Harshness)
Common Creator Pitfall
- Over-boosting highs until they become ice-pick harsh.
- On phones, excessive treble can sound cheap/tearing.
Core Approach: Gentle Low-Pass + Targeted Clarity Boost
- Use low-pass (opposite of HPF) carefully to control excessive top-end, including hiss/wireless artifacts.
- Don’t set cutoff too low or it can turn muddy.
- Suggested starting point:
- Low-pass around ~12–14 kHz with gentle slopes
“Sweet Spot” Concept
- With gentle slopes, lowering the cutoff doesn’t instantly “muffle everything” because attenuation is gradual.
Best Trick Described: High Shelf Instead of Wide High Boost
- Instead of boosting the entire high spectrum:
- Use a high shelf for subtle intelligibility lift
- Intended effect:
- Improves clarity mainly in the ~2 kHz–5 kHz intelligence region
- Reduces harsh hiss without boosting everything above.
Avoid Extreme High Shelves
- Too much boost becomes harsh and fatiguing.
De-Essing (Sibilance Control)
- Treated as dynamic EQ for sibilance.
- Warning: too much de-essing can make S’s sound like lisps or produce crunchy artifacts.
- Alternatives/workflows mentioned:
- FabFilter de-esser praised; Dyn3 called less effective
- iZotope syllable-by-syllable editing / spectral repair for gentle control
Recommendation:
- De-ess very gently—fix only what sticks out.
- If whistle-s sounds appear, spectral repair can work well (tame rather than remove everything).
Monitoring Guidance
- Speakers are preferred over headphones for EQ decisions because headphones miss the “air space” between driver and ear.
- If using headphones:
- Ensure they’re flat (calibrated/flat response), or don’t rely on them for EQ.
Consistency Rules (So Dialogue Sounds the Same Across Shots)
- Tonal consistency should come from:
- Good dialogue editing (smooth scene tone transitions)
- Avoiding too many obvious EQ changes that “call attention” to shot transitions
- General rule:
- Dialogue should remain consistent shot-to-shot to preserve performance/story connection
- Exception:
- Intentionally different scenes/lines (e.g., villain close-up delivering an epic low line) may justify more noticeable EQ/level/tone changes.
Philosophy: Avoid “Over-EQ”
- Over-EQ is a major problem.
- Mindset:
- Start flat → listen → make fewer, smaller decisions
- Each processing pass can degrade quality:
- compression, EQ, noise reduction can add artifacts
- You can’t add missing frequency content that wasn’t recorded:
- EQ can emphasize what exists, but can’t truly restore what’s absent
- Summary idea: Do less processing, do more decision-making.
Film/Music Comparison (Why Dialogue EQ Differs From Music)
- The speaker criticizes music-style “bright/icy” vocal chains used in studios (e.g., aggressive boosts over 10 kHz).
- Dialogue/video mixing should prioritize:
- natural, pleasant sound for long viewing sessions
- avoiding ear fatigue
- In film/TV/YouTube:
- avoid super shrill, highly exciter/saturated treble approaches common in some music contexts.
Plugin/Utility Announcement: Auto Level (End-of-Stream “Dessert”)
What It Does
- “Auto Level” automatically sets audio loudness to a target output level.
- It does not compress:
- it measures per-sample/level behavior and adjusts gain to match loudness.
Intended Use
- YouTube, podcasts, dialogue leveling at scale (e.g., video games with many lines)
- Live streams needing consistent spec-compliant levels
User Control
- Essentially one control: level target (input/output gain goal)
- Example targets:
- YouTube/podcast-like: around -15 output
- TV spec example: around -24 output
Technical Notes
- Includes a limiter stage (hard limiting to prevent overs)
- Latency judged low enough for lip sync tests
- CPU usage reported around ~5–12% during testing (on an M2 Ultra)
Release and Pricing Expectations
- Planned release: within 1–2 weeks
- Suggested price: around $20
- Channel members get early installer/testing access
- Likely formats discussed:
- AU and VST (AAX/Pro Tools version may be optional)
Speakers / Sources Featured
People
- Tom (primary speaker; host of the livestream)
- Ty Sonic Production (mentioned as a new channel member; not a direct speaker in the transcript)
- Other names appear as chat commenters/questions, including:
- Tiziano, Umut Bars, Lake One Sound, Kevin Broman CAS, Sasha, Chase, Simon, Greg, Brandon, Javier, Queen’s Gambit mixer, Obi, and others
Tools / Plugins / Systems Referenced
- iZotope Insight / Insight 2 (spectrogram/loudness analysis)
- iZotope Spectrum
- Pro Tools
- FabFilter Pro-Q 3
- FabFilter DNS 2
- FabFilter de-esser
- iZotope syllable editing / spectral repair
- Waves R-Bass
- Waves Saturn (specialized saturation use)
- SoundID Reference (calibration option)
- Eventide split EQ (asked about, not used)
- Pro Sub / subharmonic generator
- Pro Multiband Dynamics
- Dyn3 (Pro Tools de-esser referenced as not ideal)
- DaVinci Resolve templates/practice projects
- Pro-Q 3 / EQ3 control behaviors (as referenced during solo/band focus)
Recording / Brand References Mentioned
- Morgan Freeman, Taylor Swift, Stevie Nicks, Michael Jackson
- DJI mic (example of too-close recording)
- Rode NT1, Neumann, AKG C414
- Mentions of 32-bit recordings, field recording, and “guns” (future content)