
Farfalle is a self-hosted AI search engine for people who want a Perplexity-style search app with a choice of local or cloud models. It combines web search with model-generated answers and includes an agent that plans and carries out searches. The code is open source under Apache 2.0.
Local model support comes through Ollama, with llama3, gemma, mistral and phi3 named as supported choices. You can keep answer generation on your own hardware, or use cloud models through OpenAI and Groq. LiteLLM provides another route for connecting custom LLMs, so the app doesn't tie you to a single model provider.
Search is a separate choice. Farfalle works with SearXNG, Tavily, Serper and Bing, letting you pair your preferred search backend with your model. Running a local LLM keeps model inference local, but web search still needs an internet connection. Choosing a cloud model also sends the answer-generation work to that provider.
The app runs with Docker and presents a browser interface. Its Next.js frontend and FastAPI backend can run together on your own machine or be deployed separately. Ollama use doesn't require model-provider API keys; cloud model services and some search providers use their own keys.
For people comparing AI search apps, its main distinction is the independent choice of search provider and LLM backend. Farfalle also supports use as your browser's default search engine, so queries from the address bar can open in your self-hosted instance.
Claim this page with an email at farfalle.dev. Farfalle gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Farfalle?Promote it
Something wrong or outdated on this page?
29.8KUpdated 4 days agoApache-2.0
Docker · Web#MCP#Multi-agent workflows#Ollama integration
GPT Researcher is a self-hosted AI research agent for people who need detailed, cited reports drawn from the web, their own documents, or both. You choose the language model and search provider, and can run the software on your own server or through Docker. It's open source under Apache 2.0.
9.1KUpdated 21 hours agoMIT
macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access
Local Deep Research is a self-hosted AI research assistant for people who need cited answers drawn from academic papers, the web and their own documents. It can produce a quick summary or pursue a complex question through repeated searches, then assemble a structured report. It's open source under MIT.
9.2KUpdated 5 days agoApache-2.0
Docker · Web#Multi-user access#Ollama integration#OpenAI-compatible API
Morphic is a self-hosted AI search engine for people who want answers backed by web sources and control over the search interface they use. It combines web searches and URL reading with AI-generated responses, so you can examine the sources behind an answer. It's open source under Apache 2.0, and you can run your own instance with Docker or use the hosted website.
32.3KUpdated 24 hours ago
Docker · Web · Browser Extension#MCP#Multi-user access#Ollama integration
Onyx is an AI search and chat platform for teams whose information is spread across workplace apps. It indexes company knowledge so employees can ask questions across sources and get answers grounded in relevant documents. Teams can deploy it in their own cloud or on bare metal, including an air-gapped environment.
36.9KUpdated 4 weeks agoMIT
Docker · Web#Multimodal input#Ollama integration#OpenAI-compatible API
Vane (formerly Perplexica) is a self-hosted AI search engine for people who want answers drawn from web results, with citations they can check. It runs on your own hardware through Docker or as a server application, with a browser interface and locally stored search history. The project is free and open source under the MIT license.
18.3KUpdated 24 hours agoMIT
macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend
DocsGPT is an MIT-licensed, open-source platform for teams that want AI search, assistants and agents over their own documents. It can run on your servers with local models, including fully air-gapped deployments where documents and questions stay inside your network. Answers include the source title and page number so readers can check the evidence.