AI vocal remover & stem splitter

Remove vocals from any song.

Create karaoke tracks easily: split any song from YouTube, social links or MP3/WAV into up to 5 stems, then mix, mute and sing along with the instrumental.

About the Mazmazika vocal remover

The vocal remover separates a finished mix back into its parts. Give it a song from YouTube, SoundCloud, TikTok or Facebook, or upload an MP3 or WAV, and a neural network trained on thousands of multitrack recordings predicts which parts of the sound belong to the voice, the drums, the bass, the piano and the remaining instruments. You get each part as its own file: a clean instrumental for karaoke, an acapella for a remix, or a drum-only track for practice.

Everything runs on Mazmazika's servers, so it works on a phone as well as on a laptop, and nothing needs to be installed. A typical three-minute song is ready in under a minute with the standard engine.

How it works

  1. Paste a link or upload a file. The audio is fetched and decoded on the server; nothing is kept longer than your session needs.
  2. Choose the engine and the number of stems. Standard gives up to five stems quickly; High Quality and Ultra separate two stems, voice and instrumental, with cleaner edges and 320 kbps output.
  3. The model listens to the whole song and splits it. Progress is shown live, and the result opens in a browser mixer where each stem has its own volume and mute button.
  4. Download what you need: single stems, the instrumental, or every stem at once.

What musicians use it for

  • Karaoke and backing tracks: remove the lead vocal and keep everything else at full quality.
  • Acapellas for remixes, mashups and DJ sets.
  • Practice tracks: mute your own instrument's stem and play along with the rest of the band.
  • Transcription and study: isolate the bass or piano to hear the part clearly.
  • Vocal coaching: compare a singer's take against the isolated original.

Free vs Pro

Read live from the current plan settings, so this section always matches what the tool enforces today.

See the plans and current prices

Compared with a classic phase-cancellation "vocal cut" or a desktop stem splitter, the online remover needs no install and no GPU, and the AI separation keeps the reverb tails and harmonies that a centre-channel trick would erase.

Questions musicians ask

How good is the separation?

On well-produced pop, rock and hip-hop the instrumental is clean enough for karaoke and live use. Very reverberant recordings, live bootlegs and songs where the voice is doubled by a synth can leave faint traces. The High Quality and Ultra engines reduce those artifacts noticeably.

Which file formats can I upload?

MP3, WAV and M4A files, plus links from YouTube, SoundCloud, Facebook and TikTok. Output stems are MP3 at 128 kbps on the standard engine and 320 kbps on High Quality and Ultra.

What is a stem?

A stem is one group of instruments exported as its own audio file: vocals, drums, bass, piano, and "other" for guitars, synths and everything left. Two-stem mode gives you just voice and instrumental.

Can I use the result commercially?

The separation itself is yours, but the underlying recording keeps its copyright. Karaoke at home, practice and private study are fine; releasing a remix or performing publicly needs a licence from the rights holder.

Why is my song too long?

Each engine has a per-run track length that depends on your plan; the exact minutes are listed under Free vs Pro above and on the tool itself. The limit protects the shared queue so results stay fast for everyone.

Does it work on a phone?

Yes. The processing happens on our servers, so the page only needs a browser. Downloads land in your phone's files app, and the mixer works with touch.

Learn more