manual.raycast.comDesktop Chat Apps
#Multimodal input#Ollama integration#Tool calling
Raycast is a proprietary desktop launcher that lets you use models running through Ollama within its AI assistant. It's for people who want private, offline conversations while keeping access to Raycast's chat and AI commands. Local model requests go directly to Ollama on your computer, without sending your conversations to Raycast or third parties. Local model access is paid.
10.8KUpdated 2 days agoMIT
Browser Extension#Multilingual#Ollama integration#OpenAI-compatible API
ChatGPTBox is a free, open-source browser extension for people who want AI help with the pages they read, search results and selected text. It works in Chrome, Edge, Firefox and Safari, with mobile support too. You can connect it to Ollama or a self-hosted model through its custom model mode, or use cloud services such as ChatGPT, Claude, Moonshot and Azure.
23.8KUpdated 8 months agoMIT
Web#LLM tracing#Multi-user access#Ollama integration
Vanna is a self-hosted Python framework for building AI agents that answer database questions in plain language. It's for teams adding chat to analytics products or internal data tools, especially when each user needs different access to the data. The project is archived and no longer maintained. Its open-source code uses the MIT license.
5.7KUpdated 1 day agoApache-2.0
macOS · Windows · Linux · Web#Code execution#MCP#Ollama integration
OpenAgent is a self-hosted personal AI assistant that combines document search with agents that can act on your behalf. It's for people who want an assistant on their own computer or server, and teams building agents around their documents and workflows. It runs natively on Windows, macOS and Linux as a single executable. It's free and open source under Apache 2.0.
12KUpdated 12 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#llama.cpp backend#Multi-user access
h2oGPT is a self-hosted ChatGPT alternative for people who want to chat with local models and ask questions about their own documents. The project is archived and no longer maintained. It's open source under Apache 2.0, with support for Linux, macOS, Windows and Docker.
8.5KUpdated 1 year agoAGPL-3.0
macOS · Windows · Linux#Ollama integration#OpenAI-compatible API#RAG
Reor is a local AI note-taking app for people who want to search their own writing, find connections between ideas and ask questions about their notes. The project is archived and no longer maintained. It runs on macOS, Windows and Linux, and it's open source under AGPL-3.0.
14.2KUpdated 1 week agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
QAnything is a self-hosted knowledge base for people and teams who want to ask questions about their own documents, including collections that mix Chinese and English. It can answer in either language regardless of the document's language, and runs locally through Docker on Windows, macOS and Linux.
192Updated 1 week agoCC-BY-4.0
Linux · Docker#Multi-user access#Ollama integration
Discord Ollama Bot connects locally hosted language models to Discord, so a community can use an AI assistant inside its existing chat server. It's for server owners who want to choose and host the models themselves while keeping Discord as the interface. Model inference runs on your hardware, but conversations still pass through Discord, so this isn't a fully offline chat app.
253Updated 4 weeks agoMIT
macOS · Windows#MCP#Multilingual#Multimodal input
Raycast Ollama brings Ollama models into Raycast on macOS and Windows, for people who want AI chat and text assistance within their desktop launcher. It connects to Ollama on your own machine or a remote server you choose. The extension is open source under the MIT license, and local inference doesn't require an Ollama API key.
1.5KUpdated 2 years agoMIT
macOS · Windows · Browser Extension#Multimodal input#Ollama integration#RAG
Lumos is a Chrome extension for asking questions about web pages using models running on your own machine. It's for readers working through long discussions, product reviews, news articles or technical documentation who want summaries and answers tied to the material they're reading. Ollama handles inference locally, without a remote AI server.
1.1KUpdated 5 months agoMIT
Web · Browser Extension#Ollama integration
ollama-ui is a small browser interface for people who run models through Ollama and want a graphical place to chat. You can serve the interface locally and open it in a browser, or use the Chrome extension. It's open source under the MIT license.
10.3KUpdated 1 year agoMIT
macOS · Windows · Linux#Multimodal input#Ollama integration
Self-Operating Computer lets a vision-capable AI model control a desktop by reading the screen and choosing mouse and keyboard actions to carry out a goal. It's a Python framework for developers and researchers exploring AI agents that work through application interfaces. The code uses the MIT license.
7.7KUpdated 4 months agoBSD-3-Clause
Windows · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
Verba is a self-hosted document chatbot for people who want to ask questions across their files and knowledge bases. The project is archived and no longer maintained. It uses Weaviate to find relevant passages and gives those passages to a language model to generate answers.
4.4KUpdated 2 years agoApache-2.0
Docker · Web#Ollama integration#RAG
RAGapp is a self-hosted app for teams that want an AI assistant using retrieval-augmented generation (RAG) on their own infrastructure. It pairs a browser chat interface with an admin interface for configuring the assistant, taking an approach similar to OpenAI's custom GPTs. It's open source under the Apache 2.0 license.
12.7KUpdated 2 months agoMIT
Windows · Web#Human approval#MCP#Multi-agent workflows
Open Deep Research is a self-hosted AI agent that searches for information and writes research reports. It's for developers and teams who want to choose their own models and research tools. The project is archived and no longer maintained.
9.2KUpdated 5 days agoApache-2.0
Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API
Morphic is a self-hosted AI search engine for people who want answers backed by web sources and control over the search interface they use. It combines web searches and URL reading with AI-generated responses, so you can examine the sources behind an answer. It's open source under Apache 2.0, and you can run your own instance with Docker or use the hosted website.
3.5KUpdated 2 years agoApache-2.0
Docker · Web#Ollama integration#Web search
Farfalle is a self-hosted AI search engine for people who want a Perplexity-style search app with a choice of local or cloud models. It combines web search with model-generated answers and includes an agent that plans and carries out searches. The code is open source under Apache 2.0.
22.6KUpdated 3 months agoApache-2.0
macOS · Docker · Web#Ollama integration#OpenAI-compatible API
OpenUI is a self-hosted AI UI builder for developers prototyping web interfaces with models they choose. It turns plain-language descriptions into rendered interfaces, with a live preview and follow-up requests for changes. You can run the app locally through Docker or Python and use it in your browser.
codegpt.coCode Completion and IDE Extensions
VS Code · JetBrains#Code execution#Human approval#LM Studio integration
CodeGPT is an AI coding assistant for VS Code, JetBrains and Visual Studio that lets developers choose between models running locally through Ollama or LM Studio and cloud models such as Claude, GPT, Gemini and Grok. It's for people who want help writing and debugging code inside their editor while keeping control over which model handles each task.
kerlig.comChat With Your Documents
macOS#LM Studio integration#MCP#Multimodal input
Kerlig is a paid AI writing assistant for Mac that works with text in the apps you already use. It's aimed at people writing client emails, Slack replies and Jira tickets who want help editing and drafting without moving everything into a browser chat. It runs on Apple Silicon and Intel Macs with macOS 12 or later.
19.6KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker · Web#Ollama integration#Web search
Devika is a self-hosted AI coding agent for developers who want to give a software task in plain language and have an agent plan the work, research it and write code. Modeled after Cognition AI's Devin, it runs on your own machine with a browser interface and supports local LLMs through Ollama. It's open source under the MIT license.
hillnote.comAI Notes and Knowledge Bases
macOS · Windows · iOS · Android#MCP#Ollama integration#Works offline
Hillnote is a writing and planning workspace for people who want their notes on their own disk, with AI available inside the editor. It runs on Mac, Windows, iOS and Android. The editor, workspace and local AI work offline, and getting started doesn't require an account.
35Updated 7 months agoMIT
macOS · Web · VS Code#Code execution#LM Studio integration#MCP
Wingman-AI is a self-hosted AI agent platform for work that needs ongoing context and several agents with different roles. It suits developers and teams handling research, support, or recurring operations alongside coding tasks. A lead agent can delegate work to specialized subagents, each with its own workspace and session history.
nurgo-software.comAI Workflow Automation
Windows#Code execution#Multi-agent workflows#Multimodal input
BrainSoup is a proprietary native Windows app for people who want custom AI agents to handle work on their desktop. You can give agents distinct roles and access to different data, then have them collaborate in shared chat rooms. Natural-language conversations guide their tasks and automations.
2.1KUpdated 2 years agoMIT
macOS · Windows · VS Code#Multilingual#Ollama integration#Quantization
Llama Coder is an open source VS Code extension for developers who want a self-hosted alternative to GitHub Copilot's code completion. It uses Ollama to run models on your own hardware, either on the computer you're coding on or on a separate machine. The extension has no telemetry or tracking.
3.7KUpdated 1 day agoMIT
Docker · Web · VS Code#Git integration#Hybrid search#llama.cpp backend
Twinny is an AI coding assistant for VS Code that lets developers choose where their models run: on their own computer, a private server or a hosted API. It's for individuals and teams who want code suggestions and repository chat with control over where their code goes. The extension and team gateway are open source under the MIT license.
424Updated 7 months agoMIT
Docker#Multi-user access#Ollama integration
ollama-telegram connects Telegram chats to a local LLM through Ollama. The project is archived and no longer maintained. It's for people who want to access their own model through Telegram, with the bot and model backend running on hardware or a server they control.
14KUpdated 5 months ago
macOS · Windows · Linux · Web#GGUF#LM Studio integration#MCP
Open-LLM-VTuber is a local AI companion for people who want a character they can talk to, with a Live2D avatar that responds through speech and expressions. It runs on Windows, macOS and Linux through web and desktop clients. With local models for speech and language processing, it works fully offline and keeps conversations on your device. Cloud APIs are optional alternatives that send the corresponding processing to external services.
1.6KUpdated 1 year agoMIT
Windows · Docker · Web#llama.cpp backend#LM Studio integration#Multimodal input
Amica is a locally runnable interface for talking with customizable 3D AI characters. It's for people who want an animated, voiced character as the face of their AI assistant, with a choice of local LLM backends or cloud services. The project builds on Pixiv's ChatVRM.
9.4KUpdated 1 day agoMIT
macOS · Windows · Linux · iOS · Android#LM Studio integration#MCP#Ollama integration
Anarlog, formerly Hyprnote, is a desktop AI meeting notetaker for people who want to keep private conversations on their own hardware. It captures audio from your device without adding a bot to the call and stays hidden during screen sharing. The app runs on macOS, Windows and Linux; its community application is open source under the MIT license.
2.4KUpdated 2 years agoApache-2.0
Docker · Web#Hugging Face integration#Ollama integration
UpTrain is an open source LLM evaluation tool for developers who need to measure answer quality and investigate failures in their AI applications. Its self-hosted web dashboard runs on your machine through Docker, with a Python package for evaluations inside application code. The dashboard requires no coding.
953Updated 3 weeks agoMIT
macOS · Windows · Linux#Ollama integration
Ollama Grid Search is a desktop app for comparing LLM responses across models, prompts and inference settings. It runs on macOS, Windows and Linux, and suits developers or anyone choosing a model and prompt combination for a particular task. You can inspect the responses together rather than repeat each test by hand.
1.2KUpdated 10 months agoAGPL-3.0
Docker · Web#LLM tracing#Ollama integration
Langtrace is an open-source observability tool for developers who need to debug LLM applications and track their performance. You can run it locally or on your own servers with Docker and Docker Compose. Its traces follow OpenTelemetry standards.
6KUpdated 6 months agoMIT
Docker · Web#Multilingual#Ollama integration#OpenAI-compatible API
Paperless-AI is a self-hosted extension for Paperless-ngx users who want automatic document sorting and chat with their archive. It requires an existing Paperless-ngx instance and runs in Docker, with a browser interface for reviewing and processing documents. The project is no longer maintained.
12.3KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker#Code execution#Git integration#Human approval
ShellGPT is an AI terminal assistant for developers and people who work with shell commands. It runs on Linux, macOS and Windows, turning plain-language requests into commands suited to your operating system and shell. You can review, explain or execute its suggestions, and its Bash and Zsh integrations put generated commands into the terminal input line for editing.
1.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux#Ollama integration#RAG#Works offline
tlm is an open source terminal assistant for people who want help writing shell commands and understanding unfamiliar ones. It runs models on your workstation through Ollama, so command assistance doesn't depend on a cloud service. It works offline and requires no API key or subscription.