Demucs separates a finished song into vocals, drums, bass and the remaining accompaniment on your own computer. It's for musicians who need individual stems or a vocal-free backing track, and developers building audio tools. The project is archived and no longer maintained. Its Python code is open source under the MIT license.
The default Hybrid Transformer Demucs model combines waveform and spectrogram analysis, using a Transformer to connect the two representations of the audio. It also includes classic Hybrid Demucs and MDX models. A fine-tuned model trades longer processing time for potentially better separation, while quantized MDX models use less download and storage space with a possible loss in quality.
Demucs runs on Windows, macOS and Linux, with a community Docker option. It can process tracks on a CPU or use a CUDA GPU; typical GPU processing needs about 7 GB of GPU memory, though smaller audio segments reduce that requirement. Local processing keeps the audio on your machine. Google Colab and the Hugging Face Spaces demo offer cloud alternatives that process it remotely.
It accepts common audio formats such as WAV, MP3 and FLAC and exports separate stereo stems as WAV or MP3. Its karaoke mode produces vocals and accompaniment as two files. For people who prefer a graphical interface, Demucs-Gui and Ultimate Vocal Remover support Demucs, while a Python API lets developers include separation in their own applications.
Claim this page and we'll verify you by hand. Demucs gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Demucs?Promote it
Something wrong or outdated on this page?
1.4KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker#Batch processing#ONNX
python-audio-separator separates recordings into vocals, instrumentals and individual instruments on your own hardware. It's an open-source Python package under the MIT license, aimed at karaoke creators and developers who want audio separation in scripts or their own applications.
2.3KUpdated 10 months agoApache-2.0
macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input
62.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#Voice activity detection
26.4KUpdated 2 years agoMIT
macOS · Windows · Linux
2.6KUpdated 2 years ago
macOS · Linux · Web#Batch processing#Hugging Face integration
AudioLDM 2 generates sound effects, music and speech on your own hardware. It's a Python tool for people experimenting with synthetic audio, including sound designers and researchers who want to work with pretrained models. A Gradio browser interface and command-line tools provide access to local generation; a hosted Hugging Face demo is also available.
38.6KUpdated 2 months agoMIT
Windows · Linux · Web#Hugging Face integration#ONNX#Voice conversion
DiffRhythm is a local AI music generation model for musicians, developers and researchers who want to create full-length songs on their own hardware. It uses latent diffusion to generate songs with vocals and accompaniment, and can also produce instrumental music. The full model supports songs up to 4 minutes and 45 seconds.
GPT-SoVITS is a local text-to-speech and voice cloning tool. It can generate speech from a short reference recording or fine-tune a model for a custom voice. The source code uses the MIT license.
Ultimate Vocal Remover is a desktop app for separating vocals and other stems from audio files on your own computer. It's for people making karaoke backing tracks, isolating a vocal, or working with separate parts of a song. It runs on Windows, macOS and Linux, and its MIT license allows you to use and modify the software.
RVC WebUI is a local AI voice conversion tool for people who want to train a custom voice, change the voice in a recording, or use a live voice changer. It runs on Windows and Linux, including Ubuntu servers, with a browser interface for training and conversion and a separate interface for live use. It's free and open source under the MIT license.