7.1KUpdated 2 weeks agoMIT
Web#Code execution
LlamaCoder generates small apps from a single prompt and previews them in your browser. It's an open-source alternative to Claude Artifacts for people who want to turn an idea into an app, with a MIT license that lets them run and modify the code themselves.
55.5KUpdated 2 months ago
Docker · Web#Human approval#LLM tracing#Multi-agent workflows
Flowise is a visual builder for AI agents and chatbots that can run locally or on your own server, including through Docker. It's for developers and teams building LLM applications with connected workflow blocks. The project is archived and no longer maintained.
22.6KUpdated 3 months agoApache-2.0
macOS · Docker · Web#Ollama integration#OpenAI-compatible API
OpenUI is a self-hosted AI UI builder for developers prototyping web interfaces with models they choose. It turns plain-language descriptions into rendered interfaces, with a live preview and follow-up requests for changes. You can run the app locally through Docker or Python and use it in your browser.
33.7KUpdated 4 months ago
Windows · Docker#Code execution#Human approval#Multi-agent workflows
GPT Pilot is a self-hosted AI coding assistant for developers who want an agent to build applications under their supervision. The project is no longer maintained. It uses the FSL-1.1-MIT source-available license, with restrictions on competing uses. It runs locally as a Python CLI and is the core technology behind the Pythagora VS Code extension.
7.3KUpdated 17 hours agoApache-2.0
Linux · Docker · Web#Multi-user access
Civitai is a self-hostable platform for sharing AI image models and artwork. Its Apache-2.0 code provides accounts, model uploads, browsing and comments. The public Civitai website also operates a hosted image generator, which is separate from the runnable local platform.
55.1KUpdated 2 years agoMIT
Windows · Docker#Code execution#Multimodal input
gpt-engineer is a locally run coding assistant for developers who want to experiment with AI code generation and build their own agents. It can create software from plain-language descriptions or make requested changes to an existing codebase. The project is archived and no longer maintained. Its Python code is open source under the MIT license.
550Updated 2 years agoMPL-2.0
Web#Multilingual
Bergamot Translator is a C++ library for running neural machine translation on your own device. Developers can integrate it into native applications or build it for WebAssembly. Separate browser integrations, such as the TranslateLocally extension, make the engine accessible to end users. The software is free and open source under the Mozilla Public License 2.0.
10.8KUpdated 1 year agoMIT
Web
Zero is an open-source, self-hosted AI email app for people who want control over their inbox software while keeping accounts with providers such as Gmail and Outlook.
codegpt.coCode Completion and IDE Extensions
VS Code · JetBrains#Code execution#Human approval#LM Studio integration
CodeGPT is an AI coding assistant for VS Code, JetBrains and Visual Studio that lets developers choose between models running locally through Ollama or LM Studio and cloud models such as Claude, GPT, Gemini and Grok. It's for people who want help writing and debugging code inside their editor while keeping control over which model handles each task.
4.4KUpdated 20 hours ago
#Guardrails#Multilingual#Multimodal input
Llama Guard is Meta's collection of downloadable AI content moderation models for developers building LLM applications. It checks user inputs and model responses for content that violates safety policies, including text and images. Developers can use the models in their own deployments or access moderation through Meta's hosted Llama API.
2.9KUpdated 1 year agoAGPL-3.0
macOS · Windows · Linux · Web#Semantic search#Works offline
OpenRecall records your screen at regular intervals and makes that history searchable with local AI. It's a free, open-source alternative to Microsoft's Windows Recall and Rewind.ai for people who want to find something they previously saw on their computer. It runs on Windows, macOS and Linux, with a browser interface served from your own machine.
28.8KUpdated 4 months agoApache-2.0
Void is an open-source desktop AI code editor for developers who want to choose their own models and control where coding prompts go. It's a fork of VS Code, licensed under Apache 2.0. The project is archived and no longer maintained.
713Updated 1 year agoMIT
#RAG
PearAI is an open source AI code editor built on Microsoft’s VS Code, for developers and makers who want AI assistance inside their project. It combines questions about existing code with an agent that can write features and fix bugs. Its VS Code foundation gives it a familiar editing environment.
3.5KUpdated 4 months agoBSD-3-Clause
VS Code · JetBrains#Code execution#Persistent memory#RAG
Refact is a self-hosted AI coding assistant for developers who want an agent to work across their repository and development tools. The original project is archived and no longer maintained; ongoing development has moved to JegernOUTT/refact. The original code uses the BSD-3-Clause license.
kerlig.comChat With Your Documents
macOS#LM Studio integration#MCP#Multimodal input
Kerlig is a paid AI writing assistant for Mac that works with text in the apps you already use. It's aimed at people writing client emails, Slack replies and Jira tickets who want help editing and drafting without moving everything into a browser chat. It runs on Apple Silicon and Intel Macs with macOS 12 or later.
19.6KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker · Web#Ollama integration#Web search
Devika is a self-hosted AI coding agent for developers who want to give a software task in plain language and have an agent plan the work, research it and write code. Modeled after Cognition AI's Devin, it runs on your own machine with a browser interface and supports local LLMs through Ollama. It's open source under the MIT license.
278Updated 1 day agoAGPL-3.0
macOS · Windows · Linux · Android · Web#Works offline
Writeopia is a writing and note-taking app for people who want AI assistance while keeping their documents on their own computer. It runs on Windows, Linux and macOS, with offline access to locally stored notes and a choice of AI models that run on your device. It's aimed at writers, developers and researchers working on drafts, documentation or personal notes.
hillnote.comAI Notes and Knowledge Bases
macOS · Windows · iOS · Android#MCP#Ollama integration#Works offline
Hillnote is a writing and planning workspace for people who want their notes on their own disk, with AI available inside the editor. It runs on Mac, Windows, iOS and Android. The editor, workspace and local AI work offline, and getting started doesn't require an account.
35Updated 7 months agoMIT
macOS · Web · VS Code#Code execution#LM Studio integration#MCP
Wingman-AI is a self-hosted AI agent platform for work that needs ongoing context and several agents with different roles. It suits developers and teams handling research, support, or recurring operations alongside coding tasks. A lead agent can delegate work to specialized subagents, each with its own workspace and session history.
1.8KUpdated 2 years ago
macOS · Windows · Linux#Batch processing#Hugging Face integration
Stable Fast 3D turns a single object image into a textured 3D mesh on your own hardware. It's aimed at game and VR developers, designers and people creating product models for e-commerce. Built on TripoSR, it uses a retrained model designed to produce meshes and textures suitable for use in games and other 3D projects.
4.5KUpdated 2 years agoApache-2.0
Docker · Web#Image-to-image
InstantMesh generates a 3D mesh from a single image on your own machine. It's for creators exploring image-based 3D assets and researchers working on 3D reconstruction. The Python project is open source under Apache 2.0 and uses PyTorch with CUDA for GPU processing.
2.6KUpdated 2 years ago
macOS · Linux · Web#Batch processing#Hugging Face integration
AudioLDM 2 generates sound effects, music and speech on your own hardware. It's a Python tool for people experimenting with synthetic audio, including sound designers and researchers who want to work with pretrained models. A Gradio browser interface and command-line tools provide access to local generation; a hosted Hugging Face demo is also available.
7KUpdated 4 months agoMIT
#Batch processing
TripoSR reconstructs a 3D object from a single image on your own hardware. Developed by Tripo AI and Stability AI, it's an open-source model for researchers, developers, and artists who want image-based 3D generation they can use in their own projects. The MIT license covers both the source code and pretrained models.
1.9KUpdated 11 months ago
Docker#Batch processing#Hugging Face integration#ONNX
Nomic Embed Text v1.5 is an English text embedding model for developers building semantic search, document retrieval, and RAG applications on their own hardware or servers. It turns text into numerical representations that applications can compare by meaning. Its main distinction is adjustable embedding size: you can use smaller vectors when storage matters, with a tradeoff in retrieval quality.
4.2KUpdated 1 year agoApache-2.0
Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration
LMQL is a programming language for developers who need model calls and ordinary Python logic in the same program. It lets you define rules for generated text, including types, length limits, allowed answers and stopping phrases. Those rules apply during generation, so you can constrain intermediate responses as well as the final output.
nurgo-software.comAI Workflow Automation
Windows#Code execution#Multi-agent workflows#Multimodal input
BrainSoup is a proprietary native Windows app for people who want custom AI agents to handle work on their desktop. You can give agents distinct roles and access to different data, then have them collaborate in shared chat rooms. Natural-language conversations guide their tasks and automations.
2.1KUpdated 2 years agoMIT
macOS · Windows · VS Code#Multilingual#Ollama integration#Quantization
Llama Coder is an open source VS Code extension for developers who want a self-hosted alternative to GitHub Copilot's code completion. It uses Ollama to run models on your own hardware, either on the computer you're coding on or on a separate machine. The extension has no telemetry or tracking.
15.9KUpdated 2 months agoApache-2.0
macOS · Windows · Linux · VS Code#Git integration
DVC connects data and model versions to the code in your Git repository, so you can reproduce a machine learning experiment with the inputs it used. It's a free, open-source tool under Apache 2.0 for individual data scientists and small projects. It runs on macOS, Windows and Linux.
6.3KUpdated 9 months agoApache-2.0
Docker · Web
Aim is a free, open source ML experiment tracker for researchers and teams who want to keep training records on their own infrastructure. It runs in your training environment or on a self-hosted server, with Docker and Kubernetes deployment support. Its Apache 2.0 license permits use and modification.
3.4KUpdated 1 day agoApache-2.0
#Batch processing#Distributed execution#Multilingual
DataTrove is an open-source Python library for teams preparing large text datasets, including LLM training corpora. It runs on your own machine or on Slurm and Ray clusters, with processing steps that carry across those environments. It uses the Apache 2.0 license.
2.3KUpdated 10 months agoApache-2.0
macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input
DiffRhythm is a local AI music generation model for musicians, developers and researchers who want to create full-length songs on their own hardware. It uses latent diffusion to generate songs with vocals and accompaniment, and can also produce instrumental music. The full model supports songs up to 4 minutes and 45 seconds.
3.7KUpdated 1 day agoMIT
Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend
Twinny is an AI coding assistant for VS Code that lets developers choose where their models run: on their own computer, a private server or a hosted API. It's for individuals and teams who want code suggestions and repository chat with control over where their code goes. The extension and team gateway are open source under the MIT license.
63.5KUpdated 11 months agoMIT
macOS · Windows#Hugging Face integration
nanoGPT is a Python toolkit for developers and researchers who want to train GPT models on their own hardware or fine-tune existing GPT-2 checkpoints. Its author has deprecated the project and points readers to nanochat. The MIT-licensed code remains available for study and modification.
24.3KUpdated 5 months agoApache-2.0
VS Code#MCP#Tool calling
Roo Code is an AI coding assistant for developers working in a code editor. The project is archived and no longer maintained, and its extension has been shut down. It builds on Cline, and its source code is available under the Apache License 2.0.
3.9KUpdated 4 months agoMIT
#Batch processing#Hugging Face integration
Stable Audio Open is a text-to-audio model you can run on your own hardware to generate sound effects, field recordings and music samples. It's aimed at artists and machine learning practitioners experimenting with audio generation, and it performs better on environmental sounds and effects than on music.
5.8KUpdated 11 hours agoApache-2.0
#Hugging Face integration#Multilingual#Quantization
EmbeddingGemma is a text embedding model for developers building search and document features that run on phones, laptops or tablets. Based on Gemma 3, it converts text into numerical representations so applications can find related passages by meaning. Embeddings stay on your hardware, and the model works without an internet connection.
91Updated 2 years agoApache-2.0
#Hugging Face integration
Snowflake Arctic Embed is a family of open-source text embedding models for developers building semantic search and document retrieval systems. It turns queries and documents into numerical representations that a search system can compare by meaning. The models use the Apache 2.0 license.
huggingface.coEmbedding and Reranker Models
#Batch processing#Hugging Face integration#Multilingual
GTE (General Text Embedding) is Alibaba’s family of downloadable models for representing text as vectors. Developers use these representations to compare queries with documents, cluster related text or supply retrieval components for larger applications.
mixedbread.comEmbedding and Reranker Models
#Hugging Face integration
mxbai-embed is Mixedbread’s downloadable text embedding model family for developers building their own retrieval systems. It converts queries and passages into vectors that an application can compare for relevance or similarity.
1.8KUpdated 7 months agoApache-2.0
#Hugging Face integration
ModernBERT is a family of open-source text encoder models for developers building document search, classification and code retrieval on their own hardware. Its longer context lets it process documents and code passages that exceed the limits of older BERT models. The code and models use the Apache 2.0 license.
2.8KUpdated 1 month agoMIT
macOS#Batch processing#Hugging Face integration#LoRA
ColPali is a local AI document retrieval library for developers and researchers building document search or retrieval-augmented generation systems. It searches pages as images, using their text, charts and layout together rather than relying on a separate OCR pipeline. The colpali-engine package is deprecated; its maintainers recommend Sentence Transformers for new projects and production use.
38.1KUpdated 1 day agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Code execution#MCP#Single sign-on
Trilium Notes is a local-first note-taking app for people building a large personal knowledge base. It runs on Windows, macOS and Linux, or on your own server through Docker, with browser access and a mobile web interface. It's free and open source under AGPL-3.0.
jina.aiEmbedding and Reranker Models
Docker#GGUF#LoRA#MLX
Jina Embeddings is a family of models that converts content into vectors for retrieval, similarity matching, classification and clustering. It includes multilingual text models and multimodal variants for searching across different media.
424Updated 7 months agoMIT
Docker#Multi-user access#Ollama integration
ollama-telegram connects Telegram chats to a local LLM through Ollama. The project is archived and no longer maintained. It's for people who want to access their own model through Telegram, with the bot and model backend running on hardware or a server they control.
47.9KUpdated 1 day agoGPL-2.0
Linux · Web#LLM tracing#Multi-user access#Structured output
Discourse AI is the official AI plugin bundled with Discourse. Forum administrators can enable its features independently, including an AI bot, semantic search, topic and chat summaries, spam detection and writing assistance. It runs inside a Discourse community, which you can host on your own server.
voxta.aiAI Characters and Roleplay
Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
Voxta is an AI companion for people who want a character they can talk to, give work to or use in interactive stories. You choose its personality, voice and optional avatar. The proprietary local-server edition has a browser interface, and AI processing can run entirely on your hardware, through Voxta Cloud or across a mix of local and cloud services.
241Updated 2 years agoAGPL-3.0
Docker#Multi-user access#OpenAI-compatible API
Matrix ChatGPT Bot connects Matrix rooms to OpenAI's ChatGPT API for people who want AI conversations in their existing chat client, including Element. The project is archived and no longer maintained. It's open source under AGPL-3.0, and the maintainers point users to Baibot as an alternative.
5.1KUpdated 22 hours agoApache-2.0
Windows · Docker#Agent Skills#Hugging Face integration#ONNX
TensorRT Model Optimizer, called NVIDIA Model Optimizer or ModelOpt, is a Python library for developers preparing models for local or self-hosted inference. It reduces model size and memory use and can speed up inference through compression and other optimization techniques. It's open source under Apache 2.0.