
Locally AI is a native app for running language and vision models on recent iPhones, iPads, and Macs. It's for people who want a private AI assistant on their own device, with text, image processing, and voice conversations available without cloud processing. Once a model is downloaded, it works offline and doesn't require an account.
Model choices include Llama, Gemma, Qwen, and DeepSeek R1, alongside SmolLM, Granite, Cogito, and Liquid AI's LFM. That gives users access to models suited to conversation, multilingual tasks, reasoning, and coding within the same app. iPad and Mac users can run larger models for more demanding tasks. The app uses Apple's MLX framework and is optimized for Apple Silicon's unified memory architecture. It also integrates with on-device Apple Foundation Models.
Voice mode runs entirely on the device, so spoken conversations don't depend on a cloud service. Siri provides another way to talk to local models, while Apple Shortcuts can trigger AI actions as part of automated workflows. Control Center, the Lock Screen, and the Action Button provide quick access to the assistant.
A customizable system prompt lets you adjust the model's behavior and responses for different uses. All AI processing stays on your device, and the app states that it doesn't collect your data. Locally AI is available through the App Store for iPhone and iPad and the Mac App Store for Mac.
Claim this page with an email at locallyai.app. Locally AI gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Locally AI?Promote it
Something wrong or outdated on this page?
2.1KUpdated 8 months agoMIT
macOS · iOS#llama.cpp backend#Multimodal input#RAG
LLM Farm runs large language models offline on iOS and macOS. It's for people who want on-device AI chat or need to compare how different models perform on Apple hardware before choosing one for a project. The app is open source under the MIT license, with ggml and llama.cpp handling local inference.
2.3KUpdated 1 year agoMIT
macOS · iOS#MLX#Quantization#Works offline
noemaai.comChat With Your Documents
macOS · iOS#GGUF#MCP#MLX
3.2KUpdated 4 days agoMIT
macOS · Windows · Linux · iOS · Android#GGUF#Human approval#LM Studio integration
privatellm.appAutomation and No-Code AI
macOS · iOS#Multilingual#Quantization#Works offline
8.4KUpdated 2 days agoMIT
iOS · Android#GGUF#Hugging Face integration#llama.cpp backend
Fullmoon is an open source chat app for people who want to run language models on their Apple devices and keep conversations local. It supports iOS, iPadOS, macOS and visionOS, with on-device inference optimized for Apple silicon. You can chat fully offline, and the app saves your chat history locally.
Noema is a private local AI assistant for iPhone, iPad, Mac, and Vision Pro. It's for people who want to chat with models and work with their own files on Apple hardware without depending on a cloud service. Local chats can stay on your device, and file retrieval runs there too.
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
Private LLM runs AI chat entirely on your iPhone, iPad, or Mac. It's an app for people who want to use language models without sending their conversations to a cloud service. After the first model download, it works offline and requires no account. Conversations stay on-device, with no tracking or logs.
PocketPal AI is an open source assistant for people who want to run language models on a phone or tablet. It works on iOS, iPadOS and Android. Once you've downloaded a model, you can chat offline without an account, and your prompts, replies and documents stay on your device. The app is licensed under MIT.