Open Music and Sound Generation Models

Models that generate music, sound effects or a soundtrack for a video, including ACE-Step, YuE and Stable Audio Open.

9 tools
Local AI music separation software splits songs into stems on Windows, macOS and Linux. MIT licensed, with CPU or CUDA GPU processing.

10.4KUpdated 3 years agoMIT

macOS · Windows · Linux · Docker#Batch processing#Quantization

A local AI audio generator that turns text into sound effects, music and speech. Runs on CPU, NVIDIA CUDA or Apple Silicon with Hugging Face Diffusers support.

2.6KUpdated 2 years ago

macOS · Linux · Web#Batch processing#Hugging Face integration

An open-source AI music generator that runs locally on macOS, Windows and Linux, with text or audio style prompts and Apache 2.0 code and DiT weights.

2.3KUpdated 10 months agoApache-2.0

macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input

A local text-to-audio model for sound effects and music experiments, with CPU or CUDA support and access under the Stability AI Community License.

3.9KUpdated 4 months agoMIT

#Batch processing#Hugging Face integration

An open-source text-to-audio model that runs locally on CPU or NVIDIA GPU, with multilingual speech, voice presets and an MIT license.

39.3KUpdated 2 years agoMIT

#Hugging Face integration#Multilingual

A text-to-music model with melody conditioning and local GPU inference. AudioCraft code is MIT licensed; pretrained weights have a noncommercial license.

23.7KUpdated 2 years agoMIT

#Multimodal input

An open-source AI music model and synthesis engine that runs locally on Apple Silicon, with a macOS app and AUv3 plugin for DAWs. Apache 2.0 licensed.

1.8KUpdated 2 months agoApache-2.0

macOS#MLX#Streaming inference

Favicon of YuE

YuE

1 video
An open-source AI music generator that turns lyrics into songs. Run it locally on Linux with an NVIDIA GPU, or use the hosted demo.

10.6KUpdated 2 days agoApache-2.0

Linux · Web#Hugging Face integration#Multilingual#Multimodal input

Favicon of ACE-Step

ACE-Step

3 videos
An open-source AI music model that runs on NVIDIA GPUs and Apple Silicon, with text-to-music generation, adjustable duration and localized lyric editing.

4.9KUpdated 7 months agoApache-2.0

macOS#LoRA#Multilingual

More in Open Models