22.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration
localGPT is a self-hosted AI document chat app for people who want to question and summarise files on their own hardware. Its local Ollama setup keeps documents and conversations on your machine. Answers include source passages, so you can check what the model used.
19.2KUpdated 2 days agoGPL-3.0
macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual
pyVideoTrans translates spoken audio into another language and produces a video with translated subtitles and AI dubbing. It's for people adapting videos for audiences in other languages who want control over which parts run locally. It recognizes speech directly, so the original video doesn't need subtitles.
347Updated 2 days agoGPL-3.0
#Batch processing#LM Studio integration#Multilingual
ThunderAI brings AI writing and email processing into Thunderbird for people who want help with their inbox while choosing where their messages go. It can use local models through Ollama or an OpenAI-compatible server such as LM Studio. Cloud connections send the selected content to ChatGPT, the OpenAI API, Google Gemini or Claude instead.
76.8KUpdated 2 days agoApache-2.0
Windows#Multilingual
Tesseract is an open source OCR engine for extracting text from images, with a command line program and a library developers can embed in their own applications. It's suited to document processing workflows and software that needs text recognition. The project uses the Apache 2.0 license and doesn't include a graphical app.
superwhisper.comDictation and Voice Typing
macOS · Windows · iOS · Android#Multilingual#Works offline
Superwhisper is an AI dictation app for macOS, Windows, iOS and Android that turns speech into text in the app you're using. It's for people who prefer speaking to typing, including developers dictating requests to coding assistants. Speech recognition can run locally and offline, or use cloud models for remote processing.
1.6KUpdated 2 months agoGPL-3.0
Linux#Code execution#Multimodal input#Ollama integration
Alpaca is an open source AI chat client for people who want to run models on their own device. It uses Ollama to download and manage local models, then lets you chat without an internet connection. Conversations stay on your device in SQLite files, and local models have no direct internet access.
4.7KUpdated 5 days agoMIT
#Batch processing#Multilingual#Quantization
CTranslate2 is an open-source C++ and Python library for developers running Transformer models on their own hardware or servers. It handles translation, text generation, text encoding and speech recognition. Its custom runtime focuses on reducing inference time and memory use compared with general-purpose deep learning frameworks.
6.5KUpdated 2 months agoMIT
Linux#Multilingual#Works offline
Argos Translate runs neural machine translation locally, so you can translate text without sending it to an online service. It's for people who need offline translation and developers who want to add it to their own software. The Python library and desktop interface use the same translation engine, and the project is open source under the MIT license.
17.8KUpdated 1 day ago
Web#Git integration#Guardrails#LLM tracing
Wren AI is a self-hosted data agent for teams and agent builders who need answers based on agreed business definitions. It turns plain-language questions into SQL and interactive dashboards, using the same definitions when someone asks directly or through Claude, ChatGPT or Gemini over MCP.
25.8KUpdated 4 months agoApache-2.0
macOS · Windows · Linux · Docker · Web#Hybrid search#llama.cpp backend#Multi-user access
kotaemon is a self-hosted document chat app for people who want to ask questions across their files and check where the answers came from. It runs in a browser on Windows, macOS or Linux, with Docker also supported. The project uses the Apache 2.0 license.
11.9KUpdated 4 days agoAGPL-3.0
macOS · Windows · Linux · Android · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend
KoboldCpp pairs local model inference with a browser interface built for chat, creative writing and roleplay. A fork of llama.cpp, it bundles KoboldAI Lite with tools for keeping character details and story context alongside your conversations. It's open source under AGPL-3.0.
huggingface.coComputer Vision Models
#Hugging Face integration#Multimodal input#Structured output
Florence-2 is Microsoft's open-source vision model for developers who want to process images on their own hardware. It handles several image tasks through text prompts, so one model can generate descriptions, read text and locate objects. It runs locally with PyTorch and Hugging Face Transformers on a CPU or CUDA GPU, and uses the MIT license.
10.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
aichat brings Ollama and cloud AI services into the same terminal interface for developers and people who work at the command line. It runs locally on macOS, Linux and Windows, with Android support through Termux. Model processing happens through the backend you choose: Ollama supports local models, while providers such as OpenAI, Claude and Gemini process requests in the cloud.
21.8KUpdated 1 week agoMIT
macOS · Windows · Linux#Hugging Face integration#Multilingual#Speaker diarization
Buzz transcribes and translates speech on your own computer using OpenAI's Whisper. It's for people who need transcripts or subtitles from recordings, plus live captions from a microphone. Local transcription works offline; the optional OpenAI Whisper API sends audio to a cloud service.
395Updated 4 weeks ago
Docker · Web#Multi-user access#OpenAI-compatible API
Feeds Fun is a news reader for people whose RSS subscriptions produce more articles than they want to read. It assigns tags automatically, then uses rules you define to score articles by topic. You can self-host it with Docker or use the hosted service at feeds.fun.
12.4KUpdated 22 hours ago
iOS · Android · Docker · Web#Batch processing
Inbox Zero is an AI email assistant you can self-host with Docker or use as a hosted service. It's for people handling busy Gmail, Google Workspace, or Microsoft Outlook inboxes who want help sorting mail and preparing replies. It works alongside your existing email client.
41.9KUpdated 6 days agoGPL-3.0
macOS · Windows · Linux · iOS · Android · Web#MCP#Multimodal input#Ollama integration
Chatbox is an AI chat client for people who want local models and cloud providers in the same app. It connects to Ollama for local LLM use and supports GPT, Claude, Gemini, Grok and DeepSeek with your own API keys. Chatbox also offers its own hosted model service.
18.3KUpdated 1 day agoMIT
macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend
DocsGPT is an MIT-licensed, open-source platform for teams that want AI search, assistants and agents over their own documents. It can run on your servers with local models, including fully air-gapped deployments where documents and questions stay inside your network. Answers include the source title and page number so readers can check the evidence.
37.3KUpdated 1 day agoAGPL-3.0
macOS · Windows · Docker · Web#Batch processing#Hugging Face integration#MCP
PDFMathTranslate translates scientific PDFs while keeping their page layout, formulas, charts, contents pages and annotations. It's for researchers, students and others who need to read papers in another language without losing the relationship between the text and its figures. It produces both translated PDFs and bilingual documents for comparison with the original.
91Updated 1 day agoAGPL-3.0
Web#Multi-user access#Multilingual#Multimodal input
Nextcloud Assistant brings AI into Nextcloud Hub's documents, email, chat and calendar. It's for teams that want help with shared work while choosing where AI processing happens. With an on-premises model, data stays on your server. The app is open source under AGPL-3.0.
7.8KUpdated 2 days agoAGPL-3.0
#LM Studio integration#Multi-agent workflows#Ollama integration
Obsidian Copilot brings AI agents into your vault to find notes by meaning, edit drafts and organize research. It's for writers, researchers and other Obsidian users who want AI to work directly with their existing notes. Agents create Markdown files, wikilinks and canvases that remain in your vault.
57.6KUpdated 1 week agoApache-2.0
Docker · Web#Code execution#llama.cpp backend#MCP
PrivateGPT is a self-hosted API layer for developers building AI applications around local models. It adds document retrieval, database access and agent tools to an existing model server. Local workflows can work offline and keep data within your environment; web search and connections to online providers need internet access.
2.5KUpdated 1 day ago
macOS · Windows · Linux · Docker#Batch processing#Code execution#Multimodal input
Roboflow Inference is a self-hosted computer vision server for teams building camera and image analysis systems. It runs on your own computer, server, or edge device and combines model predictions with workflows for tracking, counting, measuring, and responding to events. Roboflow also offers hosted servers and a Serverless Cloud API, where processing runs on its infrastructure.
5.5KUpdated 6 days ago
#Semantic search#Works offline
Smart Connections is an Obsidian plugin that finds notes and passages related to what you're writing, even when you haven't linked them or entered a search query. It's for researchers, writers, and people whose vaults contain useful work they struggle to find again. Matches depend on meaning, so earlier research or a relevant decision can appear beside your current note without manual tagging.
20.3KUpdated 4 days agoApache-2.0
macOS · Windows · Linux · Docker · Web#Batch processing#Multilingual#Voice cloning
ebook2audiobook turns non-DRM ebooks into narrated audio with chapters and metadata, for readers who want audio editions of their own books. It runs locally on Windows, macOS and Linux, with Docker support and a browser interface built with Gradio. It's open source under Apache 2.0.
881Updated 3 days ago
Docker#Multilingual#Ollama integration
Lingarr is a self-hosted subtitle translator for people who maintain a media library and want subtitles in another language. It automates translation of subtitle files using a service you choose, with support for both local AI and hosted translation providers.
9.1KUpdated 22 hours agoMIT
macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access
Local Deep Research is a self-hosted AI research assistant for people who need cited answers drawn from academic papers, the web and their own documents. It can produce a quick summary or pursue a complex question through repeated searches, then assemble a structured report. It's open source under MIT.
36.2KUpdated 7 days agoMIT
#Knowledge graphs#RAG#Semantic search
GraphRAG builds a knowledge graph from text so an LLM can answer questions that depend on connections across documents or themes across a whole collection. It's for developers and researchers working with private datasets, such as business documents, proprietary research, or communications.
3.7KUpdated 5 months agoMIT
Docker#OpenAI-compatible API#Streaming inference
Speaches is a self-hosted speech server for developers who want transcription, translation and speech generation on their own hardware. Its OpenAI-compatible API lets applications use local speech models through tools and SDKs built for OpenAI's API. The project is open source under the MIT license.
23.8KUpdated 11 months ago
Docker#Code execution#RAG
PandasAI is a Python library for people who want to ask questions about their data in plain language. It works with SQL databases, CSV and parquet data, and can answer questions across multiple pandas DataFrames. Developers can use it in their own applications; analysts can use conversational queries to reduce the code they write for individual questions.
77.4KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#OpenAI-compatible API
GPT4All is a local AI chatbot for people who want to run language models on their own desktop or laptop and keep conversations on their machine. Its LocalDocs feature lets you ask questions about your own documents without sending them to a cloud service. It suits developers, teams and individuals who want control over their models and data.
1.5KUpdated 2 months agoMIT
Docker#Batch processing#Multilingual#OpenAI-compatible API
subgen generates subtitles on your own hardware for personal media libraries, including films and shows that don't have usable subtitles available. It's an open source, MIT-licensed Python service that runs in Docker or as a standalone application. Speech recognition runs locally using Whisper models through faster-whisper and stable-ts, with support for CPU processing and NVIDIA GPUs through CUDA.
40.1KUpdated 3 weeks agoApache-2.0
macOS · Linux · Web#Batch processing#llama.cpp backend#Multilingual
Marker is a local document converter for developers and teams turning PDFs, scans and Office files into structured text. It preserves tables, equations and page structure for document processing and AI workflows. Its pipeline reads embedded PDF text and uses Surya OCR where text is missing or damaged, rather than reading every page through a vision model.
29.8KUpdated 1 day ago
Docker · Web#LLM tracing#MCP#Multi-user access
FastGPT is a self-hosted AI agent builder for teams that want assistants to answer questions using company documents and carry out business workflows. Its visual editor connects model calls, knowledge retrieval and tools into applications for customer support, internal knowledge search and document review. You can run the platform on your own server through Docker or use the vendor's hosted service.
7.7KUpdated 4 weeks agoMIT
macOS · Windows · Linux#Batch processing#Multilingual#Ollama integration
Vibe is an open source desktop app for people who need transcripts or subtitles without uploading their recordings to a transcription service. It runs on macOS, Windows and Linux under the MIT license. Audio transcription works fully offline, with processing on your own computer.
31.2KUpdated 22 hours agoApache-2.0
Docker · Web#Knowledge graphs#MCP#Multi-user access
Cognee gives AI agents persistent memory across sessions, connecting documents, code, and conversations in a searchable knowledge graph. It's for developers who want agents to retain project context and teams whose knowledge sits across tickets, discussions, and repositories. The Python package is open source under Apache 2.0.