Tools tagged with "llama.cpp backend"

73 tools
A local LLM chat app for Android that runs GGUF models on-device through llama.cpp. Open source under the Apache 2.0 license.

894Updated 3 months agoApache-2.0

Android#GGUF#llama.cpp backend

An open-source React Native library that runs GGUF models on iOS and Android through llama.cpp, with GPU acceleration and image and audio understanding.

1KUpdated 3 days agoMIT

iOS · Android#GGUF#llama.cpp backend#Multilingual

An open-source LLM training toolkit that turns documents into specialist datasets, with offline generation on macOS and Linux and optional cloud compute.

1.9KUpdated 3 months agoMIT

macOS · Windows · Linux#Distributed execution#llama.cpp backend#Quantization

A Docker container build system for local AI on NVIDIA Jetson, with CUDA support and packages for Ollama, llama.cpp, PyTorch and ROS.

4.9KUpdated 3 months ago

Linux · Docker#llama.cpp backend#Multimodal input#Ollama integration

A local LLM integration for Home Assistant with voice, chat and AI automations. Runs on Raspberry Pi without a GPU and supports Ollama and llama.cpp.

1.4KUpdated 3 days ago

#GGUF#Home Assistant integration#Hugging Face integration

Local OCR model that extracts plain or formatted text from images, supports multi-page recognition, and works with Hugging Face Transformers.

8.2KUpdated 2 years ago

#Batch processing#GGUF#Hugging Face integration

An open source Android AI chat app that runs GGUF models offline through llama.cpp and connects to remote providers, including Ollama and OpenAI.

2.7KUpdated 2 weeks agoMIT

Android#GGUF#Hugging Face integration#llama.cpp backend

Local LLM desktop app for Windows, macOS and Linux. Run GGUF models offline or connect other apps through OpenAI- and Anthropic-compatible APIs.

47.7KUpdated 1 month agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#GGUF#llama.cpp backend#LoRA

A local LLM runner that packages a model and runtime in one file for macOS, Linux, BSD and Windows. Open source under Apache 2.0.

26.1KUpdated 5 hours ago

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

An open-source document chat app for Windows, macOS and Linux. Use local models through Ollama or llama-cpp-python, or connect cloud APIs.

25.8KUpdated 4 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Hybrid search#llama.cpp backend#Multi-user access

Python library for running GGUF models locally through llama.cpp, with a self-hosted OpenAI-compatible server and CPU or GPU support.

10.6KUpdated 1 week agoMIT

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

An open-source local LLM runner for Linux, macOS and Windows via WSL2, with Podman or Docker isolation and llama.cpp or vLLM inference.

3.1KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker#GGUF#Hugging Face integration#llama.cpp backend

Favicon of koboldcpp

koboldcpp

1 video
Local LLM runner for GGUF and GGML models on Windows, macOS and Linux, with CPU or GPU support, a browser UI and an AGPL-3.0 license.

11.9KUpdated 4 days agoAGPL-3.0

macOS · Windows · Linux · Android · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend

A mobile AI assistant that runs GGUF models on iOS and Android. Core chat works offline after a model download and needs no account.

8.5KUpdated 2 days agoMIT

iOS · Android#GGUF#Hugging Face integration#llama.cpp backend

A Python library that constrains LLM output with grammars and JSON schemas, supporting local Transformers and llama.cpp backends plus OpenAI.

21.8KUpdated 4 months agoMIT

#llama.cpp backend#Structured output#Tool calling

Self-hosted AI document search and agents with source citations. Run local models through Ollama, vLLM or llama.cpp, including fully offline deployments.

18.3KUpdated 1 day agoMIT

macOS · Windows · Linux · Docker · Web#Human approval#Hybrid search#llama.cpp backend

An open-source LLM vulnerability scanner that tests local Hugging Face and GGUF models or cloud APIs for security failures. Apache 2.0 licensed.

9.4KUpdated 2 weeks agoApache-2.0

#AI red teaming#GGUF#Hugging Face integration

An open-source local AI API under Apache 2.0 that connects to Ollama, llama.cpp and other OpenAI-compatible servers for document retrieval and agent workflows.

57.6KUpdated 1 week agoApache-2.0

Docker · Web#Code execution#llama.cpp backend#MCP

Favicon of Lemonade

Lemonade

1 video
An open source local AI server for chat, image generation, and speech on Windows, macOS, and Linux, with APIs for apps and agents.

5.8KUpdated 2 hours agoApache-2.0

macOS · Windows · Linux · iOS · Android · Docker#GGUF#Hugging Face integration#llama.cpp backend

Self-hosted AI research assistant for web, papers and private documents. Runs on Windows, macOS and Linux with Ollama or cloud models. MIT licensed.

9.1KUpdated 23 hours agoMIT

macOS · Windows · Linux · Docker · Web#llama.cpp backend#MCP#Multi-user access

Open-source AI framework for semantic search, RAG and agents. Runs locally or in Docker, with Hugging Face, llama.cpp and cloud models via LiteLLM.

13KUpdated 23 hours agoApache-2.0

Docker#Agent Skills#Hugging Face integration#Knowledge graphs

Self-hosted AI model serving platform for Linux, Windows and macOS. Run language, speech and image models through an OpenAI-compatible API under Apache 2.0.

9.6KUpdated 1 day agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#llama.cpp backend#Multimodal input

Favicon of llama-swap

llama-swap

1 video
A local AI proxy that switches models on demand through OpenAI and Anthropic compatible APIs. Runs on macOS, Windows, Linux and FreeBSD under MIT.

5.8KUpdated 2 days agoMIT

macOS · Windows · Linux · Docker · Web#GGUF#Image-to-image#llama.cpp backend

An open-source Android LLM chat app that runs GGUF models on-device through llama.cpp or connects to Ollama, OpenAI and Claude. Licensed under AGPL-3.0.

2.8KUpdated 1 week agoAGPL-3.0

Android#GGUF#llama.cpp backend#Ollama integration

Favicon of GPT4All

GPT4All

1 video
An open-source local AI chatbot for Windows, macOS and Linux. Run models without a GPU or cloud API, and chat privately with your documents.

77.4KUpdated 1 year agoMIT

macOS · Windows · Linux · Docker#GGUF#llama.cpp backend#OpenAI-compatible API

Local document converter turns PDFs and Office files into Markdown, JSON or HTML, with OCR on CPU, NVIDIA GPUs or Apple Silicon and optional LLM support.

40.1KUpdated 3 weeks agoApache-2.0

macOS · Linux · Web#Batch processing#llama.cpp backend#Multilingual

An MIT-licensed LLM data extraction library that validates structured outputs with Pydantic and works with Ollama, llama-cpp-python, vLLM and cloud APIs.

14KUpdated 3 weeks agoMIT

#llama.cpp backend#Ollama integration#Streaming inference

An open-source AI coding assistant for Neovim that connects to Ollama, llama-cpp or cloud providers and applies suggested edits directly to files.

18.2KUpdated 4 days agoApache-2.0

Windows#Agent Client Protocol#llama.cpp backend#Ollama integration

Favicon of Outlines

Outlines

2 videos
An open-source Python library that constrains LLM outputs to schemas and grammars, with support for local backends including Ollama and llama.cpp.

15.9KUpdated 1 month agoApache-2.0

#llama.cpp backend#MLX#Ollama integration

Local LLM library runs GGUF models through llama.cpp in Node.js, Bun and Electron. MIT licensed, with GPU support and JSON schema enforcement.

2.2KUpdated 3 days agoMIT

macOS · Windows · Linux#Batch processing#GGUF#Guardrails

An open-source LLM client that brings local models through Ollama and llama.cpp, plus cloud services such as Claude and Gemini, into Emacs.

3.5KUpdated 6 days agoGPL-3.0

#Git integration#Human approval#llama.cpp backend

Local OCR software for PDFs and images, with reading order, tables and math. Runs on CPU, Apple Silicon or NVIDIA GPUs; code uses Apache 2.0.

21.4KUpdated 3 weeks agoApache-2.0

macOS · Web#Batch processing#llama.cpp backend#Multilingual

Favicon of Msty

Msty

3 videos
Chat with local or hosted models, compare answers side by side and ask questions over your documents in Msty Studio's AI workspace.

msty.appAI Notes and Knowledge Bases

Web#llama.cpp backend#MLX#Ollama integration

Run AI models through Docker Desktop, Docker Engine or a standalone binary, with local inference and OpenAI and Ollama compatible APIs.

655Updated 2 days agoApache-2.0

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

An offline screen translator for Windows, macOS and Linux that reads Japanese game text and displays English overlays. Open source under MIT.

796Updated 2 weeks agoMIT

macOS · Windows · Linux#llama.cpp backend#LM Studio integration#Multilingual

A local AI desktop workspace for Apple Silicon Macs, Windows and Linux x64, with GGUF and MLX models, document search and optional cloud providers.

46Updated 5 days agoAGPL-3.0

macOS · Windows · Linux · Browser Extension#Code execution#GGUF#llama.cpp backend