Open-Weight AI Models to Run Locally

Open LLMs such as Qwen3 and DeepSeek, plus models for code, images, video, speech and search, with weights you can download and run yourself.

Subcategories

100+ tools
Favicon of LatentSync

LatentSync

1 video
Open-source local AI lip-sync tool with a Gradio interface and Apache 2.0 code license. GPU inference requires 8 GB or 18 GB VRAM, depending on the model.

6.1KUpdated 1 year agoApache-2.0

Web#Batch processing#Hugging Face integration#Multimodal input

Local AI 3D generator that creates meshes, radiance fields and 3D Gaussians from text or images. Runs on Linux with an NVIDIA GPU and 16GB VRAM.

13.7KUpdated 11 months agoMIT

Linux · Web#Hugging Face integration#Multimodal input

An open-weight vision model for image questions, captions and object detection. Run it locally or use hosted inference and fine-tuning.

10.1KUpdated 5 months agoApache-2.0

macOS · Windows · Linux#Hugging Face integration#Multimodal input#Works offline

Favicon of YuE

YuE

1 video
An open-source AI music generator that turns lyrics into songs. Run it locally on Linux with an NVIDIA GPU, or use the hosted demo.

10.6KUpdated 2 days agoApache-2.0

Linux · Web#Hugging Face integration#Multilingual#Multimodal input

A local LLM family for developers and researchers, with downloadable weights, text and vision models, and custom licensing for research and commercial use.

7.7KUpdated 12 months ago

#Hugging Face integration#Multimodal input#Quantization

DeepSeek-V3 and R1 are downloadable language-model families for local text generation and reasoning.

91.9KUpdated 1 year agoMIT

#Hugging Face integration

Open-source vision-language models for visual chat, document questions and image retrieval, with downloadable weights and Hugging Face Transformers support.

10.2KUpdated 1 year agoMIT

#Hugging Face integration#Multimodal input

An open source 3B local LLM that runs on CPU or CUDA GPU and supports direct or reasoning responses, six languages and a 128k context window.

3.9KUpdated 1 week agoApache-2.0

#Hugging Face integration#Multilingual#Tool calling

An open-source segmentation model for images and videos. Run it on your own GPU machine, track objects across frames, and refine masks with prompts.

19.9KUpdated 2 years agoApache-2.0

Web#Hugging Face integration

A self-hosted OCR model that extracts text and page structure from PDFs and images, with vLLM, Hugging Face Transformers and CPU inference support.

9.2KUpdated 6 months agoMIT

Docker#Hugging Face integration#Multilingual#Multimodal input

Open-source Python toolkit for semantic search and RAG, with BGE embedding models, multilingual rerankers, evaluation and fine-tuning under MIT.

12.2KUpdated 1 month agoMIT

#Multilingual#Semantic search

AI video generation model you can run on your own GPUs, with text-to-video and image-to-video support, ComfyUI integration and downloadable weights.

12.6KUpdated 3 months ago

Web#Multimodal input#Quantization

Local OCR software for PDFs and images, with reading order, tables and math. Runs on CPU, Apple Silicon or NVIDIA GPUs; code uses Apache 2.0.

21.4KUpdated 3 weeks agoApache-2.0

macOS · Web#Batch processing#llama.cpp backend#Multilingual

Favicon of PaddleOCR

PaddleOCR

2 videos
An open source OCR toolkit that runs locally, converts PDFs and images to Markdown or JSON, and supports multilingual text under Apache 2.0.

90.4KUpdated 2 weeks agoApache-2.0

Web#Multilingual#ONNX#Structured output

Open-source computer vision library for local detection, segmentation and tracking, with AGPL-3.0 licensing and exports to ONNX, TensorRT and CoreML.

62.1KUpdated 24 hours agoAGPL-3.0

#ONNX

Favicon of ACE-Step

ACE-Step

3 videos
An open-source AI music model that runs on NVIDIA GPUs and Apple Silicon, with text-to-music generation, adjustable duration and localized lyric editing.

4.9KUpdated 7 months agoApache-2.0

macOS#LoRA#Multilingual

A self-hosted AI video generation model with ComfyUI and Diffusers support, image animation, video editing, and fine-tuning tools.

11KUpdated 9 months agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input

Local image generation and editing with FLUX.1 models. Run open weights on your own infrastructure or use Black Forest Labs' hosted API.

26KUpdated 1 year agoApache-2.0

#Image-to-image#Inpainting#LoRA

Open-weight local LLMs under Apache 2.0, with 20B and 120B models that work with Ollama, LM Studio and vLLM.

20.4KUpdated 2 months agoApache-2.0

macOS · Linux#Hugging Face integration#LM Studio integration#Ollama integration

Favicon of Whisper

Whisper

6 videos
An MIT-licensed speech recognition model that runs on your own hardware, transcribes multiple languages and translates speech into English.

109.8KUpdated 4 weeks agoMIT

#Multilingual#Voice activity detection