๐Ÿ”’ Your files stay on your device โ€” core audio processing runs locally. Privacy details โ†’

How to remove vocals from a song โ€” free, in your browser.

Reduce or isolate a lead vocal for karaoke and remix sketches. Choose instant center-channel DSP or, when available, an optional UVR neural model that downloads and runs locally.

A centered vocal component reduced from a stereo music signal.
Center cancellation can reduce a centered vocal, but reverb and off-center content may remain.
๐ŸŽค Open the Vocal Remover โ†’

What "removing vocals" actually means

Many stereo mixes place the lead vocal near the center while some instruments and effects spread wider. Instant mode uses that convention: it compares left and right channels, then reduces or keeps centered vocal-frequency content. Reverb, doubled parts and anything off-center can remain.

Optional neural beta instead downloads an approximately 67 MB UVR-MDX-NET model plus its runtime and performs model-based separation locally. It is slower and requires a network download before first use, but the song itself is not uploaded. Neither mode is a magic un-mixer, so compare both and expect possible bleed or artifacts.

How to remove vocals from a song, step by step

  1. Open the Vocal Remover and drag your song onto the dropzone (or click to browse). MP3, WAV, FLAC, M4A and most other formats work. The file is decoded right there on your device โ€” you will see its length, channel count and sample rate appear once it loads.
  2. Pick a mode. Choose ๐ŸŽธ Remove vocals (instrumental) for a karaoke backing track, or ๐ŸŽค Isolate vocals (acapella) to keep only the voice. You can switch back and forth as many times as you like.
  3. Choose the engine. Instant phase DSP exposes separation intensity. Neural beta, when available, downloads and caches an approximately 67 MB model, then runs locally and may take much longer.
  4. Press โœจ Separate. Wait for the selected engine, then preview for bleed, missing instruments or artifacts. Processing time depends on track length, browser and device.
  5. Preview and download. Hit play on the Result player to A/B against the original. Happy with it? Use the download buttons to save a clean WAV, named for you as "your-track (instrumental)" or "(acapella)".

If you load a mono file, the tool flags it and falls back to a 200โ€“4000 Hz notch instead of true phase cancellation. It still works, but the result is rougher โ€” so reach for a stereo file whenever you have one.

Getting the cleanest possible result

A few habits make a big difference when you remove vocals from a song:

๐ŸŽค Try the Vocal Remover (free) โ†’

Why do it in your browser?

AudioWrench decodes and processes the selected song on-device rather than uploading it for separation. Instant DSP can work after the page assets load; neural beta requires its model and runtime download before local inference. Inputs are capped at 200 MB and decoded audio may use much more memory. For an approximate four-way heuristic decomposition, compare the Stem Splitter.

FAQ

Can you remove vocals from a song completely?
Rarely 100%. Instant DSP reduces centered vocal-frequency content and can leave reverb, doubled or off-center vocals. Optional neural beta downloads an approximately 67 MB UVR-MDX-NET model and may separate differently, but it can still leave bleed or artifacts. Preview both when available.
Is removing vocals from a song free with this tool?
Yes. There is no account or watermark. The song is processed locally, with a 200 MB input cap and device-memory limits. Neural beta also downloads an approximately 67 MB model and runtime before local processing.
Does my song get uploaded to a server?
The selected song is decoded and processed on your device rather than uploaded to AudioWrench. Neural beta does fetch its model and runtime from external hosts before local inference, so a network connection is required for that first download.
Why do I need a stereo file, and what happens with mono?
Phase-based vocal removal works by comparing the left and right channels and cancelling the shared center, so it needs a true stereo mix. If you load a mono file the tool tells you and switches to a 200โ€“4000 Hz notch fallback that simply ducks the vocal frequency range. The fallback is rougher and dents the instruments a little, so feed it stereo whenever you can.

Sources and review notes

Instant center-channel DSP and optional neural separation are different modes with different speed, download and artifact trade-offs.

Keep reading

Related tools