Open-Weight AI Models to Run Locally

Open LLMs such as Qwen3 and DeepSeek, plus models for code, images, video, speech and search, with weights you can download and run yourself.

Subcategories

100+ tools
Favicon of Chatterbox

Chatterbox

3 videos
An open-source text-to-speech model family that runs on your own hardware, clones voices from short clips, and supports offline deployment.

26.6KUpdated 2 months agoMIT

Linux#Multilingual#Voice cloning#Voice conversion

An open-weight local LLM family from Google DeepMind for phones, PCs and servers, with Ollama and LM Studio support and an Apache 2.0 JAX library.

5.8KUpdated 2 days agoApache-2.0

Android#LM Studio integration#LoRA#Multilingual

Favicon of IndexTTS

IndexTTS

1 video
Local text-to-speech software clones voices from one audio clip, supports five languages, and provides separate controls for emotion and speaking speed.

24.2KUpdated 1 day ago

Windows · Linux · Web#Hugging Face integration#Multilingual#Multimodal input

Favicon of Wan2.2

Wan2.2

3 videos
Open source video generation models that run on your GPU, with text, image, speech and character animation options under Apache 2.0.

17.7KUpdated 1 week agoApache-2.0

#Hugging Face integration

Self-hosted text-to-speech with voice cloning, multilingual speech and emotion control. Code and weights use the FISH AUDIO RESEARCH LICENSE.

32.9KUpdated 2 weeks ago

#Batch processing#Multilingual#Multimodal input

Favicon of Qwen3

Qwen3

1 video
A language model family with public weights for local CPU or GPU use through Ollama, llama.cpp, and LM Studio, plus server deployment.

27.7KUpdated 9 months ago

#Batch processing#GGUF#Hugging Face integration

Favicon of VibeVoice

VibeVoice

2 videos
Open-source voice AI models for local transcription and speech generation, with MIT licensing, CPU inference, and streaming audio support.

54.5KUpdated 4 weeks agoMIT

#Hugging Face integration#Multilingual#Quantization

Local AI music separation software splits songs into stems on Windows, macOS and Linux. MIT licensed, with CPU or CUDA GPU processing.

10.4KUpdated 3 years agoMIT

macOS · Windows · Linux · Docker#Batch processing#Quantization

Text embedding models for self-hosted multilingual and code search, with custom vector sizes and support for CPU or NVIDIA GPU inference.

2KUpdated 1 year ago

Docker#Batch processing#Hugging Face integration#Multilingual

Python toolkit for semantic search and RAG with BGE embedding models, multilingual rerankers, fine-tuning and evaluation. MIT licensed.

12.2KUpdated 1 month agoMIT

#Multilingual#Multimodal input#Semantic search

AI content moderation models check text and images in LLM inputs and outputs. Use downloadable models or Meta's hosted Llama API.

4.4KUpdated 20 hours ago

#Guardrails#Multilingual#Multimodal input

A local AI audio generator that turns text into sound effects, music and speech. Runs on CPU, NVIDIA CUDA or Apple Silicon with Hugging Face Diffusers support.

2.6KUpdated 2 years ago

macOS · Linux · Web#Batch processing#Hugging Face integration

An open-source image-to-3D model that runs locally with Python and PyTorch. MIT-licensed code and weights, with about 6GB VRAM for default inference.

7KUpdated 4 months agoMIT

#Batch processing

A local text embedding model for semantic search and RAG, with adjustable vector sizes, Apache 2.0 licensing, and support for Sentence Transformers.

1.9KUpdated 11 months ago

Docker#Batch processing#Hugging Face integration#ONNX

An open-source AI music generator that runs locally on macOS, Windows and Linux, with text or audio style prompts and Apache 2.0 code and DiT weights.

2.3KUpdated 10 months agoApache-2.0

macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input

A local text-to-audio model for sound effects and music experiments, with CPU or CUDA support and access under the Stability AI Community License.

3.9KUpdated 4 months agoMIT

#Batch processing#Hugging Face integration

A multilingual text embedding model that runs offline on phones, laptops and tablets, with open weights and a quantized memory footprint under 200MB.

5.8KUpdated 2 days agoApache-2.0

#Hugging Face integration#Multilingual#Quantization

Text embedding models for semantic search, licensed under Apache 2.0, with compact, large and long-context variants for document retrieval.

91Updated 2 years agoApache-2.0

#Hugging Face integration

Alibaba’s GTE models turn text into vectors for retrieval and similarity matching, with downloadable weights and local Python inference.

huggingface.coEmbedding and Reranker Models

#Batch processing#Hugging Face integration#Multilingual

Mixedbread’s mxbai-embed models produce text vectors locally for document retrieval, with Apache-2.0 weights and adjustable embedding dimensions.

mixedbread.comEmbedding and Reranker Models

#Hugging Face integration

Open-source text encoder models under Apache 2.0 for local retrieval and classification, with long context and Hugging Face Transformers support.

1.8KUpdated 7 months agoApache-2.0

#Hugging Face integration

A local document retrieval library that matches text queries to page images without OCR. MIT-licensed Python code supports NVIDIA and Apple Silicon GPUs.

2.8KUpdated 1 month agoMIT

macOS#Batch processing#Hugging Face integration#LoRA

Jina’s embedding models encode multilingual text and media for retrieval, with local weights, noncommercial licenses and commercial deployment options.

jina.aiEmbedding and Reranker Models

Docker#GGUF#LoRA#MLX

Higgs Audio V2, now Higgs TTS 2, is a downloadable speech model for expressive narration, multilingual dialogue and voice cloning.

8.4KUpdated 4 months agoApache-2.0

#Batch processing#Hugging Face integration#Multilingual

Text embedding models for local retrieval, with English and multilingual variants, Hugging Face checkpoints, and MIT-licensed Python code.

22.2KUpdated 1 week agoMIT

#Hugging Face integration#Multilingual

Open-source text-to-speech software that runs locally, generates English speech and supports voice cloning. MIT licensed, with downloadable models.

4.7KUpdated 1 year agoMIT

#Hugging Face integration#Multilingual#Voice cloning

Open-source text-to-speech model with voice cloning. Runs locally on Linux and macOS under Apache 2.0, with a hosted audio playground also available.

7.2KUpdated 2 years agoApache-2.0

macOS · Linux · Docker · Web#Multilingual#Voice cloning

An open-source text-to-speech model you can run locally, with MIT-licensed Python code, pretrained English voices and adaptation to unfamiliar speakers.

6.4KUpdated 3 years agoMIT

Windows#Hugging Face integration#Multilingual#Voice cloning

An open-source speech generation model that uses text and audio context, runs on a CUDA-compatible GPU, and integrates with Hugging Face Transformers.

14.7KUpdated 1 year agoApache-2.0

Windows#Hugging Face integration#Multimodal input

Open-source AI video generator that runs locally with miniFLUX or SD3 models, supports Apple Silicon, and includes a browser interface.

3.2KUpdated 2 years agoMIT

macOS · Web#Hugging Face integration#Multimodal input

An open-source audio AI model for local transcription, translation and Q&A. Runs offline with vLLM or Transformers under Apache 2.0.

10.8KUpdated 3 months agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

Favicon of XTTS v2

XTTS v2

1 video
Local text-to-speech model with voice cloning and streaming audio, available through Coqui TTS on Linux, macOS and Windows.

2.3KUpdated 4 months agoMPL-2.0

macOS · Windows · Linux · Docker#Multilingual#Streaming inference#Voice cloning

An open-source AI animation module for Stable Diffusion that runs locally, works with personalized models, and supports image or sketch guidance.

12.3KUpdated 2 years agoApache-2.0

Web#Hugging Face integration#LoRA#Multimodal input

An image-to-video model for animating still pictures locally, with Python and browser demos. Model weights use Stability AI Community terms.

27.3KUpdated 9 months agoMIT

Web#Hugging Face integration

Open source AI video generator you can run on your own GPUs, with image and text inputs, model training tools, and an Apache 2.0 license.

29.9KUpdated 6 months agoApache-2.0

#Hugging Face integration#Multimodal input

Favicon of SkyReels

SkyReels

1 video
AI video generator with a hosted web app for synchronized sound and reference control, plus downloadable models for local GPU inference.

7.6KUpdated 8 months ago

Web#Hugging Face integration#Multimodal input