Audio Editor · Windows 10 and 11
Get the acapella out of a song
An acapella is the voice of a song with nothing behind it. Octoolo isolates it with AI that runs on your own computer: separate the song, press Voice only, and export the singing as a WAV, FLAC or MP3.
How to isolate the vocals from a song
- 1
Add the song
Drop the song into the Audio Editor and click it on the timeline.
- 2
Separate it
Press Separate voice and music on the right, then Separate. The first time, Octoolo downloads the AI part once and shows its size first.
- 3
Keep only the voice
Under Parts of the song, press Voice only. Play it through and listen to the quiet parts.
- 4
Export a lossless file
Press Export and choose WAV or FLAC for a remix or a sampler, or MP3 to listen.
Lined up, for remixes
The voice comes out exactly as long as the song and starts at the same moment, so it lines up with the original's beat. Export the whole length rather than trimming it, and it will sit on the grid of any music program at the song's tempo. Octoolo keeps the parts as 44.1 kHz FLAC, which is what Demucs, the model it uses, works at.
To keep a short phrase as a sample instead, split the voice at the start and end of the line with Split and delete the rest, then export. Fade the edges by a tenth of a second so it does not click.
How clean is an extracted voice?
The voice is pulled out by a model trained to tell singing from instruments, so it is very good, not perfect. What to listen for:
- Bleed. On loud drum hits or cymbal crashes a faint ghost of them can stay in the voice.
- Effects. Reverb and delay on the original voice come along with it, since they are part of how it was mixed.
- Harmonies. Backing vocals stay in the voice track. They cannot be separated from the lead.
Lowering the voice part's volume a little and adding Reduce background noise can soften a hissy trace. For the cleanest acapellas, start from the best file you have: a WAV or FLAC from the original release separates better than a low-bitrate MP3 or a video soundtrack.
Voices from speech, too
The model treats speech as voice, so the same steps lift a speaker out of a video with music behind it, or an interview with a song playing in the room. Separate the clip's sound, keep Voice only, and the music drops away. Other people talking stay in the voice track, since they are voices too.
Octoolo or an online tool?
| Detail | Octoolo | A typical online tool |
|---|---|---|
| Your files | Stay on your computer | Usually uploaded to the site's server |
| Without internet | Works | Does not work |
| File size | No limit | Often capped on free use |
| Watermarks and ads | None | Common on free plans |
| Cost | $4.99 a month, all 17 apps | Free with limits, or a plan per site |
Questions, answered
Which format should I export the acapella in?
WAV or FLAC for remixing, sampling or another editor, since both keep every detail. MP3 is fine for listening or sending.
Will the acapella line up with the original song?
Yes. It is as long as the song and starts at the same moment, so it stays in time with the original and its tempo.
Can it remove the backing vocals and keep only the lead?
No. All singing goes to the voice track together.
Is anything uploaded?
No. Octoolo separates the song on your computer. Only the AI model is downloaded, once.
Acapella extractor, and 16 more apps.
Download for Windows7 days free, then from $3.99 a month for all 17 apps. Windows 10 and 11.