What the vocal remover does not do
This is centre-channel cancellation, not a source-separation model, and it has real limits worth knowing before you rely on it.
Needs a centred, dry vocal
It works best on older or simply mixed recordings. A modern mix that widens the lead vocal, doubles it, or adds heavy reverb will not cancel cleanly, leaving a quieter ghost of the voice rather than silence.
Cancels more than the voice
The technique has no idea what a voice is. Anything else placed dead centre, such as the snare or a centred lead instrument, is subtracted along with it. Only content below 120 Hz is protected.
Not a clean acapella
The isolated vocal file is the centre of the mix above 120 Hz, which includes the voice plus anything else centred. It is useful for following a lyric or a melody, not as a clean vocal stem.
Weaker in mono playback
A phone speaker or a mono PA adds the left and right channels back together, which brings the cancelled content back and weakens the instrumental track, especially at a high intensity setting.
Not a source-separation model
A trained source-separation model handles modern, reverb-heavy mixes far better and can produce real instrument stems. Those models need much heavier processing than a browser tab, and most web versions of them upload your audio to a server to run it.
Related tools
Frequently asked questions
Yes. It cancels the signal shared between the left and right channels, so a mono file, which has no sides to compare, is rejected with an error instead of processing silently.
No. Vocal cut intensity only changes how much of the centre is subtracted from the instrumental track. The isolated vocal file is always the full centre signal above 120 Hz, wherever the slider is set.
No. It is centre-channel cancellation, arithmetic that runs with the Web Audio API. There is no model, no training data and nothing to download beyond the page.
Often only partly. It does best on older or simply mixed recordings with a centred, dry vocal. A mix that widens or doubles the lead vocal, or adds a lot of reverb, leaves a ghost of the voice instead of silence.
Because everything centred above 120 Hz is removed, not only the voice. A centred snare or lead instrument goes with it. Lowering the intensity leaves more of the mix in, at the cost of leaving more vocal too.
No. Decoding, the centre-channel arithmetic and the WAV encoding all run in this tab with the Web Audio API. No library is downloaded to do it.
Your files stay on your device
Your browser decodes the song and runs the centre-channel arithmetic inside this tab. Nothing is uploaded, and no library is downloaded either: separation uses only the browser's built-in Web Audio API.