Acapella extractor: isolate the vocals from any song in your browser

Drop in a track and SplitVocals pulls the voice out as its own file — an acapella you can remix, sample, study or lay over a new beat — using an AI model that runs entirely on your device, with no upload. HD Pack and Pro add a cleaner server-side vocal.

The 4-stem browser separation carries no charge. Server-side HD, 6 stems, batches and FLAC are paid — see pricing.

Extract the acapella

or drop it here · MP3, WAV, FLAC, M4A, OGG · up to 50 MB / 10 min

Browser separation runs locally · Uploads only if you choose HD

How a finished mix gives up its acapella

There is no vocal track hidden inside an MP3 waiting to be found; the voice was mixed together with everything else. SplitVocals runs Demucs v4, a hybrid-transformer model trained on thousands of songs to recognise what a singing voice sounds like inside a full arrangement, and reconstructs it as a separate stream. The lead vocal, the harmonies and the ad-libs all count as voice and come out together in the vocal stem; drums, bass and the rest of the band land in their own stems.

What you get is a standard audio file — WAV at the source sample rate or 320k MP3 — that drops straight into a DAW, a sampler or a DJ deck. The other three stems come from the same pass, so the instrumental is one click away if you need it too.

Who pulls acapellas

Producers and remixers lay a studio vocal over a new beat, retime it, or chop it into a hook. DJs blend an acapella over another record's instrumental for a live mashup. Singers and vocal coaches study phrasing, breath and riffs with the band out of the way, then sing along to the isolated line. Songwriters and transcribers hear a lyric or a harmony part that the mix buried. Sound designers and video editors pull a spoken or sung line to use as a texture. None of these needs an account; the separation runs in the browser tab.

What makes an acapella clean, and what makes it dirty

A clean acapella starts with a clean recording. Studio vocals that were tracked on their own microphone separate almost completely; live recordings, where the singer bleeds into every other mic on stage, do not. Reverb and delay on the voice are part of the voice — the model keeps the effect with the vocal, and there is no way to get a drier take than the one that was recorded. Instruments that live in the same range as the voice, such as a lead guitar or a synth lead doubling the melody, can leave faint traces in the stem.

Give the model the best source you have: a lossless WAV or FLAC, or a streaming-quality file rather than a low-bitrate rip. If a vocal still comes back with bleed, the HD Pack and Pro run a larger model on a server that separates vocals noticeably more cleanly, and Pro adds a six-stem split that also isolates guitar and piano.

Unreleased material stays on your machine

Producers often need an acapella from a demo, a session bounce or a vocal that has not been released. On this page the file is decoded and processed inside your browser through WebGPU or WebAssembly and is never uploaded — there is no server copy to leak, expire or sit in someone's queue. Open the Network tab of your developer tools during a separation and you will see no outgoing request with your audio in it.

It is also faster: no upload of a large WAV, and once the ~170 MB model is cached after the first run, every separation starts instantly, even offline. A typical song takes one to three minutes on Chrome or Edge with WebGPU; Safari and Firefox use a slower CPU path.

How it works

  1. 1

    Drop a song

    MP3, WAV, FLAC, M4A, OGG or AAC, up to 50 MB or 10 minutes.

  2. 2

    AI splits it on your device

    A Demucs v4 model runs locally via WebGPU or WebAssembly — nothing is uploaded.

  3. 3

    Play, solo, download

    Preview each stem, then export WAV, 320k MP3, or a ZIP of all four.

Why running locally matters

Private by construction

Your audio is decoded and processed inside the browser tab. There's no upload step, so there's nothing to verify in a privacy policy — check the Network tab yourself.

No queue, no waiting

Processing starts the moment you drop a file. There's no server-side render queue, so a slow day for other users never slows you down.

Works offline after first use

The AI model downloads once (~170 MB) and is cached by your browser. Come back tomorrow, even without a connection, and separation still works.

No per-song cost

There is no per-song server compute to pay for in the browser tier, so nothing is metered. Server-side HD is the paid exception, and it says so before you click.

What is included, and what is paid.

The 4-stem browser separation on this page carries no charge and no cap on songs. The HD Pack and Pro exist for the parts that need a server: studio-grade HD quality, more stems, batches and lossless files.

Browser separation

$0

  • 4-stem separation, on your device, no cap on songs
  • No account, no upload, no queue
  • MP3 and WAV downloads, no watermark
  • Supported by ads

HD Pack and Pro

HD Pack $4.99 one-time · Pro $5.99/month or $39/year

  • Server-side HD vocals and instrumental, cleaner than a browser can manage
  • 6 stems, adding guitar and piano (Pro)
  • Batches of 5 files and lossless FLAC export (Pro)
  • No ads (Pro)
See pricing

Frequently asked questions

How do I extract the acapella from a song?
How do I extract the acapella from a song?
Drop the file into the box above. SplitVocals separates the song on your device and offers the isolated vocal for download as WAV or 320k MP3, alongside the instrumental and the individual drum, bass and other stems.
Does the acapella include backing vocals and harmonies?
Does the acapella include backing vocals and harmonies?
Yes. The model recognises every sung voice as vocal, so the lead, the harmony stack and any ad-libs come out together in the vocal stem. Separating a lead singer from their backing vocalists is not something the model does.
Can I remove the reverb from the extracted vocal?
Can I remove the reverb from the extracted vocal?
No. Reverb and delay were printed onto the vocal when the song was mixed, so the model treats them as part of the voice and keeps them with it. A drier source recording gives a drier acapella; there is no setting that dries a wet one.
Is extracting an acapella paid?
Is extracting an acapella paid?
No — the acapella on this page comes from a separation that runs on your own device, so there is no per-song charge and no daily cap. The paid HD Pack and Pro add server-side HD vocals, six stems, batches and FLAC export; see the pricing page for what each costs.
What format does the acapella come in?
What format does the acapella come in?
Lossless WAV at the source sample rate, or MP3 at 320k. Pro adds FLAC export. The stem is stereo and the same length as the original, so it lines up sample-for-sample with the instrumental from the same pass.
Can I release a remix built on an extracted acapella?
Can I release a remix built on an extracted acapella?
Only with the rights to the original recording. Extracting a vocal for practice, study or a private mashup is one thing; releasing it needs clearance from whoever owns the master and the song, the same as sampling would. SplitVocals adds no restriction of its own on top of that.