
Fullmoon is an open source chat app for people who want to run language models on their Apple devices and keep conversations local. It supports iOS, iPadOS, macOS and visionOS, with on-device inference optimized for Apple silicon. You can chat fully offline, and the app saves your chat history locally.
Its model choices include Llama 3.2 1B and 3B, plus DeepSeek-R1-Distill-Qwen-1.5B. Fullmoon supports 4-bit versions of the Llama models and both 4-bit and 8-bit versions of the DeepSeek model. These smaller models make it an option for people looking for local AI chat on an iPhone or iPad as well as a Mac.
Fullmoon also connects local models to Shortcuts. You can use a shortcut to chat with a model or pass its output to other actions, so the app can take part in automation beyond its own chat interface. Appearance settings let you change the theme and fonts, while a custom system prompt lets you adjust how the model responds.
The app uses Swift MLX and Metal for model processing on Apple hardware. It runs inference on your device rather than sending prompts to a cloud model service. Fullmoon's Swift source code is available under the MIT license.
Claim this page with an email at fullmoon.app. Fullmoon gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Fullmoon?Promote it
Something wrong or outdated on this page?
locallyai.appDesktop Chat Apps
macOS · iOS#MLX#Multilingual#Multimodal input
Locally AI is a native app for running language and vision models on recent iPhones, iPads, and Macs. It's for people who want a private AI assistant on their own device, with text, image processing, and voice conversations available without cloud processing. Once a model is downloaded, it works offline and doesn't require an account.
noemaai.comChat With Your Documents
macOS · iOS#GGUF#MCP#MLX
privatellm.appAutomation and No-Code AI
macOS · iOS#Multilingual#Quantization#Works offline
2.1KUpdated 8 months agoMIT
macOS · iOS#llama.cpp backend#Multimodal input#RAG
3.2KUpdated 4 days agoMIT
macOS · Windows · Linux · iOS · Android#GGUF#Human approval#LM Studio integration
6KUpdated 3 months agoApache-2.0
macOS · iOS#Multimodal input#Ollama integration#Works offline
Noema is a private local AI assistant for iPhone, iPad, Mac, and Vision Pro. It's for people who want to chat with models and work with their own files on Apple hardware without depending on a cloud service. Local chats can stay on your device, and file retrieval runs there too.
Private LLM runs AI chat entirely on your iPhone, iPad, or Mac. It's an app for people who want to use language models without sending their conversations to a cloud service. After the first model download, it works offline and requires no account. Conversations stay on-device, with no tracking or logs.
LLM Farm runs large language models offline on iOS and macOS. It's for people who want on-device AI chat or need to compare how different models perform on Apple hardware before choosing one for a project. The app is open source under the MIT license, with ggml and llama.cpp handling local inference.
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
Enchanted is a ChatGPT alternative for macOS, iOS and visionOS that connects to models on your own Ollama server. It's for people who want an Apple app for chatting with privately hosted models such as Llama 2, Mistral, Vicuna and Starling. You supply the model server. The app is open source under the Apache License 2.0.