
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
The app supports GGUF models such as Llama, Qwen 3, Gemma 3, Phi-4 and Mistral. Inference uses the CPU or a supported GPU, including Metal on Apple Silicon and Adreno GPUs on Android. If a model is too large for your phone, you can connect to Ollama, LM Studio, LocalAI or another OpenAI-compatible server on your local network. The model then runs on that server.
Stable Diffusion generates images on-device. Vision models can answer questions about photos, and Whisper transcribes spoken input locally. You can attach PDFs, CSV files and code to a conversation, or search saved documents through a local knowledge base.
Downloaded local models work offline. Web search and connected services need network access, and using a remote model sends requests to the server you select.
Claim this page with an email at getoffgridai.co. Off Grid AI gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Off Grid AI?Promote it
Something wrong or outdated on this page?
noemaai.comChat With Your Documents
macOS · iOS#GGUF#MCP#MLX
Noema is a private local AI assistant for iPhone, iPad, Mac, and Vision Pro. It's for people who want to chat with models and work with their own files on Apple hardware without depending on a cloud service. Local chats can stay on your device, and file retrieval runs there too.
66.6KUpdated 2 hours agoMIT
macOS · Windows · Linux · Android · Docker · Web#LM Studio integration#MCP#Multimodal input
41.9KUpdated 6 days agoGPL-3.0
macOS · Windows · Linux · iOS · Android · Web#MCP#Multimodal input#Ollama integration
52.3KUpdated 1 hour agoAGPL-3.0
macOS · Windows · Linux#LM Studio integration#MCP#Multilingual
390.8KUpdated 19 hours ago
macOS · Windows · Linux · iOS · Android · Docker#Multi-user access#Persistent memory#Tool calling
layla-network.aiAI Characters and Roleplay
iOS · Android#Code execution#GGUF#llama.cpp backend
AnythingLLM is an open source AI assistant for people who want to chat with their documents and use AI agents on their own hardware. The desktop app runs on Windows, macOS and Linux, while a Docker deployment supports multiple users on a self-hosted server. No account is required for the desktop app.
Chatbox is an AI chat client for people who want local models and cloud providers in the same app. It connects to Ollama for local LLM use and supports GPT, Claude, Gemini, Grok and DeepSeek with your own API keys. Chatbox also offers its own hosted model service.
Cherry Studio is a free, open source desktop app for people who use several AI models and want their conversations in one place. It runs on Windows, macOS and Linux. Chats, settings and knowledge base files are stored on your device.
OpenClaw is a self-hosted AI assistant that connects your own computer to the messaging services you already use. It's for people who want a personal assistant on their laptop, or teams that want to run a shared assistant on their own hardware. The project is open source under the MIT license.
Layla is an offline AI assistant for Android and iOS that runs language models on your phone. It's for people who want private everyday chat or an AI companion with custom characters, memory and roleplay. Local conversations stay encrypted on your device; optional cloud mode sends requests to your chosen hosted provider.