
Ollama runs language models on your own computer or server. It provides a command-line runner and a local API for people building AI applications or connecting existing tools to models they host themselves. The software is distributed under the MIT license.
It runs on macOS, Windows and Linux, with a Docker image for server deployments. You can download models from its library or import your own, then run inference locally. Hardware requirements depend on the model you choose.
Applications can connect through Ollama's REST API or its Python and JavaScript libraries. Coding agents, chat interfaces and automation tools can use the same local runtime, so you can change the model behind a workflow without having to build a model server yourself.
Claim this page with an email at ollama.com. Ollama gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Ollama?Promote it
Something wrong or outdated on this page?
153.6KUpdated 1 week ago
Docker · Web#Code execution#Human approval#Hybrid search
Open WebUI gives individuals and teams a self-hosted place to chat with local LLMs and cloud models. It runs on your own computer or server, including through Docker, and can work entirely offline with local models.
lmstudio.aiComputer and Browser Agents
macOS · Windows · Linux#llama.cpp backend#MCP#MLX
130KUpdated 39 minutes agoMIT
Web#Code execution#GGUF#Hugging Face integration
37.8KUpdated 1 day agoAGPL-3.0
Docker · Web#Web search
655Updated 2 days agoApache-2.0
macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend
77.4KUpdated 1 year agoMIT
macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#OpenAI-compatible API
LM Studio is a desktop application for downloading and running language models on macOS, Windows and Linux. You can search for models, manage downloads and chat with them through the app. Downloaded models can run offline, including document chat that uses files on your computer.
llama.cpp runs language models on your own hardware and can serve them from a machine you control. It’s an MIT-licensed, open source inference engine for people building local AI apps, running a private model server, or using a model directly from the command line. It supports vision-language models too.
SearXNG is a self-hosted metasearch engine for people who want web search without user tracking or profiling. It brings results from separate search services into one browser interface. You can run your own instance or use a public one, depending on who you want to trust with your searches.
Docker Model Runner lets developers run and serve AI models on their own computer or server using Docker Desktop, Docker Engine or the standalone dmr binary. It pulls models from Docker Hub, OCI registries, and Hugging Face, then stores them locally. Inference runs locally too.
GPT4All is a local AI chatbot for people who want to run language models on their own desktop or laptop and keep conversations on their machine. Its LocalDocs feature lets you ask questions about your own documents without sending them to a cloud service. It suits developers, teams and individuals who want control over their models and data.