Remove Background Noise from Audio

Clean hiss, hum and room noise out of a recording with an FFT denoiser at three strengths, plus an optional rumble cut. Runs in your browser.

🌐 Español

Drop your file here (.mp3, .wav, .m4a, .ogg, .flac)

🔒 Private by design: your files are processed locally in your browser and never uploaded to any server.

What a denoiser can see, and what it cannot

Noise reduction works on a statistical assumption: that unwanted sound is steady and wanted sound is not. Hiss from a cheap preamp, the tone of an air conditioner, the hum of mains electricity, the general roar of a room, all of these hold roughly the same shape across the whole recording. A voice does not.

So the filter builds a picture of the steady component in each frequency band and subtracts it, leaving what varies behind. That works remarkably well on hiss and hum, and it does nothing at all about a door slamming, a chair creaking or someone talking in the next room, because none of those are steady.

Understanding that boundary is the difference between a satisfied user and a frustrated one. If your problem is a constant background tone, this will help a great deal. If your problem is intermittent, no amount of turning the strength up will fix it, and turning it up will damage what is left.

Cleaning up an interview recording

  1. Drop your file on the box or use Choose a file. MP3, WAV, M4A, OGG and FLAC are accepted, one at a time.
  2. Set Noise reduction strength. Medium (recommended, 12 dB) is where to start.
  3. Leave Also cut low rumble and hum (below 80 Hz) ticked for speech. Untick it for music with real bass.
  4. Click Remove Background Noise from Audio. The progress bar tracks the encode.
  5. Download the result, which is named after your file with a denoised suffix and an mp3 extension.
  6. Listen on headphones before you commit to it. Denoising artifacts hide well on laptop speakers.
  7. Process another clears the box.

If medium leaves too much noise, try strong before reaching for anything else, but compare the two directly rather than assuming more is better.

The three strengths, in decibels

The strength selector is not a vague slider. Each option maps to a specific amount of noise reduction expressed in decibels, which is how far the filter pushes the detected noise floor down.

Light is 6 dB. That is a gentle lift, enough to take the edge off a faint hiss without touching the character of the voice. It is the safe choice for material you plan to master properly later.

Medium is 12 dB, and it is the underlying filter’s own default for good reason. On a typical laptop or phone recording it makes a clearly audible difference and almost never damages speech.

Strong is 24 dB. On a genuinely noisy recording it can be the difference between usable and unusable. It also introduces the classic aggressive-denoising sound: a faint watery warble in the quiet gaps, sometimes described as musical noise, caused by isolated frequency bins surviving the subtraction. Use it when the alternative is discarding the recording.

Why the rumble cut is on by default

Below 80 Hz, a spoken voice carries almost nothing. A typical male speaking voice bottoms out a little above that line and a female voice considerably higher.

What does live down there is everything you did not want: traffic passing outside, the thrum of an air conditioner, the low thump of a hand brushing the microphone or the desk, and the fundamental of mains hum at 50 Hz in Europe or 60 Hz in North America. Cutting the band costs speech nothing and removes a surprising amount of muddiness.

It runs first, before the denoiser, so the denoiser is not asked to model energy that is being discarded a moment later.

The one case to untick it is music. A bass guitar’s low E is around 41 Hz and a kick drum’s weight sits below 80 Hz, so the cut would hollow out the track.

Where denoising sits in a workflow

Denoise before you level, not after. Normalize Audio will bring the result to a consistent loudness, and doing that first would simply amplify the noise you are about to remove. The same argument applies to Boost MP3 Volume.

If the recording has long dead patches, Remove Silence from Audio will tighten it up afterwards. For a transcript, Transcribe Audio works better on a cleaned file than a noisy one, so run this first. And if the problem is that you need the vocal separated from a backing track rather than cleaned, that is a different job entirely and Vocal Remover is the tool for it. The rest is on the audio tools hub.

See it in action

Screenshot of the Remove Background Noise from Audio tool with sysfenix-sample.wav (517 KB) loaded, Noise reduction strength set to Medium (recommended, 12 dB), Also cut low rumble and hum (below 80 Hz) set to on
Remove Background Noise from Audio mid-process: sysfenix-sample.wav (517 KB) loaded, Noise reduction strength set to Medium (recommended, 12 dB), Also cut low rumble and hum (below 80 Hz) set to on.
Screenshot of the Remove Background Noise from Audio result screen showing sysfenix-sample-denoised.mp3 ready to download (72 KB, 86% smaller)
The finished result: sysfenix-sample-denoised.mp3 ready to download (72 KB, 86% smaller). The download link is a local blob URL — the file never leaves your device.

Frequently asked questions

Why does the file always come back as an MP3?

Because denoising rewrites every sample, so there is no way to avoid re-encoding, and once a re-encode is unavoidable one predictable output format is better than four. Everything comes out as MP3 at 192 kbps regardless of whether you fed in a WAV, an M4A, an OGG or a FLAC. If your source was lossless, keep the original, because this output is not.

Which strength should I use?

Medium unless the recording tells you otherwise. Light applies a gentle 6 dB reduction and suits material with a faint hiss you want tidied rather than removed. Medium is 12 dB and is the underlying filter's own default. Strong pushes to 24 dB and will visibly help a genuinely noisy recording while risking the watery, warbling artifacts that aggressive noise reduction is known for.

What does the rumble option remove?

Everything below 80 Hz, applied before the denoiser runs. That band holds traffic, air conditioning, handling noise and the fundamental of mains hum at both 50 and 60 Hz, and it holds almost nothing of a speaking voice, whose lowest useful energy sits a little above it. For speech this is close to free. For music with real bass content, untick it.

Does it need a sample of pure noise first?

No, and that is a deliberate design choice. Studio denoisers usually ask you to select a couple of seconds of silence so they can profile the noise floor. This uses continuous noise tracking instead, so the filter estimates the floor as it goes. That suits a one click tool, and it also copes better with a recording where the noise changes partway through.

Can it remove someone talking in the background?

No. The filter targets steady, broadband noise such as hiss, hum and room tone, which look statistically flat over time. Speech is not steady, so a second voice is treated as signal rather than noise and survives the process largely intact. Nothing available in a browser will separate two overlapping speakers.

Will it fix a recording that clipped?

No. Clipping is destroyed information, not added noise. The waveform was flattened at the moment of recording and the peaks that were cut off are simply not in the file. Denoising cannot invent them back, and running it on a clipped recording usually just makes the distortion easier to hear.

Does the order of the two filters matter?

Yes, and it is fixed. The rumble cut runs first so the denoiser never has to model energy in a band that is about to be thrown away, which leaves it a cleaner picture of the noise you actually care about. Running them the other way around would waste part of the denoiser's work on frequencies destined for the bin.

How long can the recording be?

There is no hard limit written into this tool, but there is a practical one. Everything runs on a single threaded engine inside your tab, so a one hour recording takes a long time and consumes real memory. For a long interview, cut it into pieces first, denoise each, and rejoin them afterwards.

Related tools