Lumos is a Chrome extension for asking questions about web pages using models running on your own machine. It's for readers working through long discussions, product reviews, news articles or technical documentation who want summaries and answers tied to the material they're reading. Ollama handles inference locally, without a remote AI server.
The extension uses retrieval-augmented generation (RAG) to bring page content into the conversation. You can ask about a whole page or narrow the context to text you've highlighted. Site-specific content extraction lets it focus on relevant sections, such as an article body or forum comments.
Documents work too. Lumos accepts PDFs, CSV and JSON files, along with plain text formats such as Markdown and source code. You can also use clipboard text as an attachment. An attached file becomes the conversation's source material in place of the current page.
With a multimodal model, Lumos can include images from the page in your prompts or process attached JPEG and PNG files. It also lets you save and reopen chats, regenerate the last response and cancel a request.
Lumos is open source under the MIT license. The browser extension depends on a local Ollama server for its model and embedding workflow; the model doesn't run inside Chrome itself. You can choose an Ollama language model such as llama2 and an embedding model such as nomic-embed-text. The local backend can also run in Docker.
Claim this page and we'll verify you by hand. Lumos gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Lumos?Promote it
Something wrong or outdated on this page?
46Updated 5 days agoAGPL-3.0
macOS · Windows · Linux · Browser Extension#Code execution#GGUF#llama.cpp backend
Vyact is an open source desktop workspace for people who want to use a local LLM with their documents, email and code. It runs on Apple Silicon Macs, Windows and Linux x64 under the AGPL-3.0 license. Intel Macs aren't supported.
23.8KUpdated 19 hours agoMPL-2.0
macOS · Windows · Linux · iOS · Android#Multilingual#Persistent memory#RAG
1.1KUpdated 7 months agoAGPL-3.0
Web · Browser Extension#Multilingual#Multimodal input#Ollama integration
5.4KUpdated 3 months ago
macOS · Windows · Linux#MCP#Multilingual#Ollama integration
5ire is a free desktop AI assistant that combines chat, a local document knowledge base and MCP tools. It runs on macOS, Windows and Linux, with Mac downloads for Apple Silicon and Intel. It's for people who want to use their own documents and external tools alongside conversations with local or cloud models.
10.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
66.6KUpdated 2 hours agoMIT
macOS · Windows · Linux · Android · Docker · Web#LM Studio integration#MCP#Multimodal input
Brave Leo is an AI assistant built into the Brave browser, with Bring Your Own Model support for people who want to use their own local or remote models while browsing. It can work with third-party APIs as well as Brave's hosted model choices. The browser runs on macOS, Windows, Linux, Android, and iOS.
NativeMind brings a local AI assistant into Chrome for people who want help with webpages, documents, and writing without sending that content to a cloud model. The browser extension connects to Ollama on your machine, keeping prompts and AI processing on-device. It requires no account and uses the AGPL-3.0 open-source license.
aichat brings Ollama and cloud AI services into the same terminal interface for developers and people who work at the command line. It runs locally on macOS, Linux and Windows, with Android support through Termux. Model processing happens through the backend you choose: Ollama supports local models, while providers such as OpenAI, Claude and Gemini process requests in the cloud.
AnythingLLM is an open source AI assistant for people who want to chat with their documents and use AI agents on their own hardware. The desktop app runs on Windows, macOS and Linux, while a Docker deployment supports multiple users on a self-hosted server. No account is required for the desktop app.