Free · No account · Unlimited splits

Split any song into isolated stems

Upload a track and get separate files for vocals, drums, bass, piano and everything else — ready to preview in the browser and download.

Useful for karaoke backing tracks, remixes, transcription, practice loops and rebalancing a mix you can no longer re-open.

Your file is uploaded to our processing server, separated, and then deleted — nothing is stored permanently and no account is needed. See the privacy policy for details.

What VocaSplitter does

VocaSplitter takes a finished stereo recording and reconstructs the individual parts inside it as separate audio files. Upload one song and you get back a vocal track with no band behind it, a drum track with nothing else on it, a bass line on its own, and so on — each as a normal audio file you can play, edit or import into any DAW.

It runs a source separation model of the Spleeter family on our server. That matters for one practical reason: separation is a genuinely heavy computation, and doing it server-side means the quality does not depend on how fast your laptop or phone is. A three-minute song typically takes one to three minutes depending on how busy the queue is.

The tool is free and has no upload limit, no sign-up and no watermark on the output. It is built and maintained independently, and running costs are covered by advertising on the written guides rather than by charging for separations.

How it works, in three steps

01

Upload your track

Drop in an MP3, WAV, M4A, OGG or FLAC file. Nothing is required beyond the file itself — no account, no email, no watermark on the result.

02

Pick a separation mode

Two stems for a clean vocal-and-instrumental split, three to also pull the drums out, or five for vocals, drums, bass, piano and everything else.

03

Preview, then download

Play each stem in the browser before committing to anything. Download individual files, or take the whole set as a single ZIP.

For the detail behind each step — what the model is actually doing, and why the source file you choose matters more than any setting — read how VocaSplitter works.

Choosing a separation mode

More stems is not the same as better quality. Every additional stem is another decision the model has to get right, so the mode with the fewest classes that answers your question will usually sound best.

ModeFiles you getBest for
2 stemsVocals, InstrumentalKaraoke tracks, acapellas, removing a voiceover
3 stemsVocals, Drums, InstrumentalDrum practice, beat transcription, DJ edits
5 stemsVocals, Drums, Bass, Piano, OtherRemixing, rebalancing a mix, learning one instrument

Read the full decision guide for how source material changes the answer.

What people use it for

Karaoke and backing tracks

A two-stem split gives you an instrumental you can sing over. Our karaoke guide covers the level and ghosting work that makes a track usable in a real room.

Remixing and DJ edits

Loop an isolated drum stem to extend an intro, drop the beat for a breakdown, or build a mashup from an acapella and a separate instrumental.

Learning parts by ear

Hearing a bass line without the kick drum on top removes the hardest step in transcription. Mute your own instrument and play against the original rhythm section.

Teaching and rehearsal

Hand each section of an ensemble a reference of their own part, or show a student the difference between what is on the record and what they are playing.

Video and podcast editing

Pull a music bed out from under a voiceover, or recover a clean narration track from a mix where the session files are long gone.

Rebalancing a finished mix

When the multitrack is lost, separated stems let you bring a buried vocal forward or pull down a piano that sits too loud, then bounce a new version.

What to realistically expect

Separation is an estimate, not a recovery of the original session. When a song was mixed down, dozens of sources were summed into two channels and the information about what came from where was genuinely lost. The model reconstructs a very well-informed guess. Understanding that explains almost everything you will hear.

Sparse recordings separate beautifully — a voice with a guitar, a soul record, most pre-2000 rock. Dense, heavily-limited modern productions with layered synths are much harder, and you will hear a faint shadow of the vocal in the instrumental. Songs with heavy autotune or a hard-doubled lead are harder still, because the voice no longer resembles the natural singing the model learned from.

Source quality matters more than any setting. A 128 kbps MP3 has already had most content above roughly 16 kHz discarded by the encoder, and the model will faithfully separate a song with no top end. Feed it a WAV, a FLAC, or at minimum a 256 kbps MP3.

For practice, transcription, karaoke, DJ edits and sample sourcing, the quality is well past the point of being useful. For commercial release of an isolated vocal it usually is not — you can hear the process on headphones if you listen for it. Our guide to fixing common artifacts covers what can and cannot be repaired afterwards.

Advertisement

Supported files

  • MP3, WAV, M4A, AAC, OGG, FLAC and AIFF.
  • Stereo or mono. Stereo separates better, since the model can use position as a cue.
  • Any sample rate. Output comes back at the sample rate you uploaded.
  • If a file is rejected, re-export it as a 44.1 kHz WAV — that repairs most container-level problems.

Your files and your privacy

Uploads are sent to our processing server over HTTPS, separated, and then deleted. We do not keep a library of what you upload, we do not train models on it, and we do not pass it to anyone else.

Because files are removed after processing, download your stems during the session rather than bookmarking the links. Full detail is in the privacy policy.

Guides and tutorials

Longer write-ups on how separation works and how to get real work out of it.

Technology · 9 min read

How AI Stem Separation Actually Works

A plain-English explanation of how a neural network pulls vocals, drums and bass out of a finished stereo mix — spectrograms, masking, and why artifacts happen.

Tutorials · 8 min read

How to Make a Karaoke Track From Any Song

A complete walkthrough for producing a clean, singable backing track — choosing the right split mode, dealing with leftover vocal ghosting, and getting the level right for a live room.

Guides · 6 min read

2, 3 or 5 Stems: Choosing the Right Separation Mode

More stems is not better quality. A decision guide covering what each separation mode actually gives you, when the extra classes help, and when they quietly destroy your audio.

Troubleshooting · 9 min read

Fixing the Most Common Stem Separation Artifacts

Watery vocals, hi-hat bleed, missing bass punch and vocal ghosting — what causes each artifact and the specific repair that actually works.

Tutorials · 7 min read

Getting Separated Stems Into Your DAW Without Losing Sync

Sample rates, alignment, tempo detection and gain staging — the practical steps for turning downloaded stems into a session you can actually work in.

Guides · 8 min read

Stem Separation and Copyright: What You Can and Cannot Do

Separating a song creates a derivative work. A practical, non-legalistic look at private use, covers, remixes, sampling, karaoke and monetised video — and where the lines actually fall.

Guides · 7 min read

How Musicians Use Isolated Stems to Learn Songs Faster

Practical practice methods built on separated audio — transcribing a buried bass line, drilling a groove against the original drummer, and building ear training that actually transfers.

Please only process audio you have the right to use

Splitting a commercial recording creates a derivative work. Doing that for private study or practice is uncontroversial in most places; publishing, performing publicly, or monetising the result generally needs permission from whoever controls the recording and the composition.

We explain where the lines fall in stem separation and copyright, and the terms of use set out what we ask of you.