Vocal splitter: split any song into vocals, drums, bass and more — in your browser

SplitVocals is an AI vocal splitter: it separates a song into four stems (vocals, drums, bass, other) with a model that runs entirely on your device — no upload and no account for the browser separation. HD Pack and Pro add server-side quality.

The 4-stem browser separation carries no charge. Server-side HD, 6 stems, batches and FLAC are paid — see pricing.

Split into 4 stems

or drop it here · MP3, WAV, FLAC, M4A, OGG · up to 50 MB / 10 min

Browser separation runs locally · Uploads only if you choose HD

One model, four stems, zero uploads

Most "free" stem splitters send your file to a server, put you in a queue, and cap you at a handful of minutes per day. SplitVocals works differently: it downloads a Demucs v4 hybrid-transformer model straight into your browser and runs the separation on your own CPU or GPU. Your audio file is read locally and never crosses the network — you can confirm that yourself by opening your browser's DevTools Network tab while a song processes.

Because nothing touches a server, there's no per-minute quota, no watermark, and no reason to ask for an email address. Drop in a track, get vocals, drums, bass and "other" (guitars, keys, synths, everything else) back as separate files, with no per-song charge in the browser tier.

Built for the people who actually need stems

DJs and producers pull individual stems to build mashups and remixes without re-recording anything. Cover singers and karaoke hosts strip the lead vocal to get an instrumental in one pass. Drummers, bassists and other musicians mute their own instrument to practice against a real backing track instead of a static play-along. Teachers isolate a single part to demonstrate phrasing or technique. None of that requires a subscription; Pro adds server-side HD, six stems and batches for people who want more. Your output is yours, for personal and commercial use.

Getting a clean separation

Separation quality tracks the quality of the source mix. A well-produced studio recording with clearly panned, distinct instruments splits cleanly. Heavily compressed, reverb-drenched, or live-recorded material — where instruments physically bleed into each other's microphones — can leave faint traces of one stem inside another. If a result sounds off, try the original studio or streaming-quality file rather than a lossy re-encode; more bits in gives the model more to work with.

The first time you separate a track, your browser downloads the ~170 MB Demucs model and caches it. Every separation after that — even offline — skips the download and starts immediately.

Why running locally actually matters

Local processing isn't just a privacy nicety. It means no upload time for large files, no wait in a server queue behind other users, and no service outage stopping you mid-project. Chrome and Edge use WebGPU to run the model on your graphics card, finishing a typical song in roughly one to three minutes on a laptop; Safari and Firefox fall back to a slower CPU path via WebAssembly. Either way, the work happens on hardware you already own.

How it works

  1. 1

    Drop a song

    MP3, WAV, FLAC, M4A, OGG or AAC, up to 50 MB or 10 minutes.

  2. 2

    AI splits it on your device

    A Demucs v4 model runs locally via WebGPU or WebAssembly — nothing is uploaded.

  3. 3

    Play, solo, download

    Preview each stem, then export WAV, 320k MP3, or a ZIP of all four.

Why running locally matters

Private by construction

Your audio is decoded and processed inside the browser tab. There's no upload step, so there's nothing to verify in a privacy policy — check the Network tab yourself.

No queue, no waiting

Processing starts the moment you drop a file. There's no server-side render queue, so a slow day for other users never slows you down.

Works offline after first use

The AI model downloads once (~170 MB) and is cached by your browser. Come back tomorrow, even without a connection, and separation still works.

No per-song cost

There is no per-song server compute to pay for in the browser tier, so nothing is metered. Server-side HD is the paid exception, and it says so before you click.

What is included, and what is paid.

The 4-stem browser separation on this page carries no charge and no cap on songs. The HD Pack and Pro exist for the parts that need a server: studio-grade HD quality, more stems, batches and lossless files.

Browser separation

$0

  • 4-stem separation, on your device, no cap on songs
  • No account, no upload, no queue
  • MP3 and WAV downloads, no watermark
  • Supported by ads

HD Pack and Pro

HD Pack $4.99 one-time · Pro $5.99/month or $39/year

  • Server-side HD vocals and instrumental, cleaner than a browser can manage
  • 6 stems, adding guitar and piano (Pro)
  • Batches of 5 files and lossless FLAC export (Pro)
  • No ads (Pro)
See pricing

SplitVocals vs. typical cloud stem splitters

 SplitVocalsTypical cloud tool
Audio upload requiredNeverUsually required
Processing limitsUnlimited songsMinutes/credits per day or plan
Account / signupNot neededOften required
CostFreeFree tier + paid plans
Where processing happensYour deviceRemote servers
Wait timeStarts immediatelyUpload + render queue

Frequently asked questions

What is included at no charge, and what is paid?
What is included at no charge, and what is paid?
Included: the 4-stem browser separator. It runs on your device, costs us nothing per run, and has no cap on songs. Paid, and optional: the HD Pack ($4.99 one-time, 10 server-side HD separations) and Pro ($5.99 a month or $39 a year) for HD quality, six stems, batches of five and FLAC export.
Do I need to sign up or install anything?
Do I need to sign up or install anything?
No account and no download. Open the page in Chrome, Edge, Safari, or Firefox and drop in a file. The AI model is fetched automatically the first time you separate a track and then cached for future use.
Does my audio get uploaded anywhere?
Does my audio get uploaded anywhere?
No. The file is decoded and processed inside your browser tab using WebGPU or WebAssembly. You can verify this yourself by opening your browser's Network tab during a separation — you'll see no outgoing request carrying your audio.
What stems do I get?
What stems do I get?
Four: vocals, drums, bass, and "other" (guitars, keys, synths, and anything else that isn't one of the first three). You can download each stem individually, as a ZIP of all four, or as a ready-made mixdown like an instrumental or acapella.
What file types and sizes are supported?
What file types and sizes are supported?
MP3, WAV, FLAC, M4A, OGG, AAC and WebM, up to 50 MB or 10 minutes per file. If a track is longer, trim it first or split it into sections.
Why is my first separation slow?
Why is my first separation slow?
The very first run on a new device downloads the ~170 MB AI model to your browser's cache. Every separation after that, including future visits and offline use, starts right away without re-downloading.