Audio Silence Detection, RMS Gate Trimming & Dead Air Truncation
Audio Silence Detector & Trimmer automatically detects and removes dead air, pauses, and silent intervals from podcast recordings, voiceover auditions, lecture audio, and acoustic stems. Featuring customizable decibel threshold gating (-60 dB to -20 dB), minimum silence duration detection, and smooth crossfade padding, it condenses spoken audio tracks and tightens speech rhythm without creating unnatural, clipped transitions.
Manually hunting down and cutting dozens of silent gaps in a two-hour podcast episode wastes hours of editing time. This automated silence cutter operates 100% client-side on decoded Web Audio PCM buffers, delivering optimized, edited audio tracks in seconds without cloud uploads.
コアアーキテクチャ&計算式
RMS Energy Calculation: \text{RMS}(k) = \sqrt{\frac{1}{W} \sum_{i=0}^{W-1} x^2[k \cdot W + i]} ; \text{Threshold (dBFS)} = 20 \log_{10}(\text{RMS})
Segments audio into 20ms analysis windows, computes Root Mean Square (RMS) energy, flags continuous sub-threshold blocks, and splices non-silent segments with 25ms crossfades.
ベストプラクティスと必須ガイドライン
- Calibrate Noise Threshold Against Ambient Room Tone: Set your silence gate threshold 6 to 10 dB above the baseline ambient room noise floor (typically between -42 dB and -36 dB for standard room acoustics). Setting the threshold too high cuts into quiet speech syllables and trailing word endings.
- Maintain 150ms to 250ms Buffer Padding to Preserve Speech Cadence: Completely eliminating all pauses results in breathless, robotic speech delivery. Adding 150-250 milliseconds of pre- and post-speech padding preserves natural conversational breathing and natural rhythm.
- Apply Micro-Crossfades Between Spliced Audio Segments: Abruptly joining separated audio chunks creates high-frequency voltage clicks at transition points. Always ensure smooth 15-30ms crossfades are applied between joined audio boundaries to maintain seamless room tone.
- Process Multi-Speaker Podcasts Separately Before Merging: If podcast hosts are recorded on separate microphone tracks, run silence removal on each stem individually before combining them. This eliminates ambient room bleed and coughs on idle microphones.