ChatterUI is an Android chat app for people who want to run a local LLM on their phone or use the same interface with a remote model. It supports assistant conversations and character chats, with controls for how chats are structured and how models generate replies. It's open source under AGPL-3.0.
Local Mode runs GGUF models on the device through llama.cpp. The model must fit in your phone's memory, and inference happens on the phone rather than through a remote API. ChatterUI can use model files already in device storage or keep its own copy. Android is the available platform; iOS isn't available.
Remote Mode connects the mobile app to backends such as Ollama, koboldcpp and text-generation-webui, including those you host yourself. It also supports cloud services such as OpenAI, Claude and Cohere, alongside Open Router, Mancer and AI Horde. In this mode, the selected backend processes chat requests. Generic text and chat completion support, plus custom API templates, extend the app beyond its named connections.
For character conversations, ChatterUI supports Character Card v2 and keeps multiple chats per character. You can adjust generation settings and the instruction format rather than rely on a fixed chat setup. It also works with the phone's text-to-speech engine to read responses aloud.
Claim this page and we'll verify you by hand. ChatterUI gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find ChatterUI?Promote it
Something wrong or outdated on this page?
8.4KUpdated 2 days agoMIT
iOS · Android#GGUF#Hugging Face integration#llama.cpp backend
PocketPal AI is an open source assistant for people who want to run language models on a phone or tablet. It works on iOS, iPadOS and Android. Once you've downloaded a model, you can chat offline without an account, and your prompts, replies and documents stay on your device. The app is licensed under MIT.
layla-network.aiAI Characters and Roleplay
iOS · Android#Code execution#GGUF#llama.cpp backend
2.7KUpdated 2 weeks agoMIT
Android#GGUF#Hugging Face integration#llama.cpp backend
3.2KUpdated 4 days agoMIT
macOS · Windows · Linux · iOS · Android#GGUF#Human approval#LM Studio integration
894Updated 3 months agoApache-2.0
Android#GGUF#llama.cpp backend
16.2KUpdated 1 day agoApache-2.0
Windows · iOS · Android#Image-to-image#Multimodal input#ONNX
Layla is an offline AI assistant for Android and iOS that runs language models on your phone. It's for people who want private everyday chat or an AI companion with custom characters, memory and roleplay. Local conversations stay encrypted on your device; optional cloud mode sends requests to your chosen hosted provider.
Maid is an Android AI chat app for people who want to run models on their phone and access remote models in the same app. It runs GGUF models locally through llama.cpp without an internet connection. It's open source under the MIT license, with no ads or telemetry.
Off Grid AI runs language models on iOS, Android, macOS, Windows and Linux. You can chat, analyze documents and generate images on your own hardware. The mobile app uses the MIT license, while the desktop app uses AGPL.
SmolChat is an Android app for people who want to chat with language models running on their own phone. It uses llama.cpp to run GGUF models on-device, including small language models. Inference stays on your device.
MNN is a lightweight C++ framework for developers who want AI models to run on phones, PCs and embedded devices. It handles inference and training on the device, with a focus on small application footprints and hardware acceleration. The project is open source under Apache 2.0, and Alibaba uses it in apps including Taobao, Youku and DingTalk.