Tools tagged with "Tool calling"

200+ tools
Python RAG framework for document-based AI apps, with local models through Ollama, cloud APIs, and customizable retrieval workflows.

39.6KUpdated 1 year ago

#Ollama integration#RAG#Reranking

Desktop AI assistant for macOS, Windows and Linux that connects to Ollama or cloud providers, with a local document knowledge base and MCP tools.

5.4KUpdated 3 months ago

macOS · Windows · Linux#MCP#Multilingual#Ollama integration

Favicon of AutoGen

AutoGen

2 videos
AI agent framework for Python with local and distributed runtimes, a browser-based prototyping UI, and OpenAI and Azure OpenAI integrations.

61.2KUpdated 6 months agoCC-BY-4.0

Web#Code execution#MCP#Multi-agent workflows

An open-source AI agent framework that runs locally in Docker, supports local LLMs with a GPU, and provides a browser interface for managing agents.

17.7KUpdated 2 years agoMIT

Docker · Web#Human approval#Persistent memory#Tool calling

Self-hosted AI application server with an OpenAI-compatible API, local Ollama and vLLM backends, document search and agent tool calling. MIT licensed.

8.4KUpdated 21 hours agoMIT

#Agent Skills#Batch processing#Guardrails

An offline AI assistant for Android and iOS that runs models on your phone, keeps chats encrypted on-device, and offers optional hosted models.

layla-network.aiAI Characters and Roleplay

iOS · Android#Code execution#GGUF#llama.cpp backend

Local LLM web interface for Windows, macOS and Linux. Use GGUF models, Ollama or cloud APIs, with local chat storage and an Apache 2.0 license.

4.8KUpdated 3 weeks agoApache-2.0

macOS · Windows · Linux · Docker · Web#GGUF#Hugging Face integration#llama.cpp backend

Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

An on-device AI SDK that runs text, image and audio models on macOS, Windows and Linux, with GGUF, MLX and an OpenAI-compatible API.

qualcomm/GenieXInference Libraries and Bindings

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

A local vision-language model for image analysis, document extraction and video understanding, with Apache 2.0 weights and Hugging Face Transformers support.

huggingface.coOCR and Document Scanning

Linux#Batch processing#Hugging Face integration#Multimodal input

Multimodal AI models you can run locally with Mistral's GPU inference library, which is open source under Apache 2.0 and archived.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Multimodal input#Tool calling

Self-hosted AI data assistant that queries databases, analyzes files, and generates reports with local models or cloud APIs. Open source under MIT.

20.1KUpdated 2 days agoMIT

macOS · Linux · Docker · Web#Code execution#llama.cpp backend#OpenAI-compatible API

Self-hosted AI agent framework in C# for Windows, Linux and macOS, with LLamaSharp and cloud provider plugins. Apache 2.0 licensed.

3.1KUpdated 2 days agoApache-2.0

macOS · Windows · Linux · Web#Code execution#MCP#Multi-agent workflows

Downloadable coding models for local completion and tool-using agents. Codestral and Devstral releases have different model licenses.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Tool calling

An open weights LLM you can self-host with Transformers, with document citations, tool calling and a CC-BY-NC-4.0 noncommercial license.

huggingface.coOpen-Weight LLMs

#Guardrails#Hugging Face integration#Multilingual

A family of AI models you can run offline with Ollama, llama.cpp or LM Studio, with open weights and training data for building specialized agents.

2.1KUpdated 3 weeks agoApache-2.0

Linux#GGUF#Guardrails#Hugging Face integration

LLM models for local or cloud inference, with downloadable weights and function calling. The Apache 2.0 inference library is archived.

10.8KUpdated 3 months agoApache-2.0

Docker#Hugging Face integration#Multimodal input#Tool calling

Open-source LLM for self-hosted AI agents, with thinking and direct-response modes, MIT licensing, and support for vLLM and SGLang.

huggingface.coCoding Models

#Hugging Face integration#LoRA#Multilingual

An open-source Go library for building LLM applications with Ollama, OpenAI and Gemini, plus components for agents and document retrieval. MIT licensed.

9.7KUpdated 9 months agoMIT

#Ollama integration#RAG#Semantic search

A self-hostable AI chat interface with Docker deployment, local model connections through LocalAI and RWKV-Runner, and desktop clients.

88.8KUpdated 2 months agoMIT

macOS · Windows · Linux · Docker · Web#MCP#Multimodal input#OpenAI-compatible API

A self-hosted language model series for coding and tool use, with Base and Instruct variants and a hosted OpenAI/Anthropic-compatible API.

11.1KUpdated 11 months ago

#Hugging Face integration#Quantization#Tool calling

A self-hosted language model for reasoning and agent tasks, with fast and slow thinking modes, a 256K context window, and Transformers and vLLM support.

820Updated 1 year ago

#Quantization#Tool calling

Language models for English, Korean and Spanish, with reasoning and tool use. Includes an on-device model and GGUF, GPTQ and AWQ formats.

107Updated 1 year ago

#GGUF#Hugging Face integration#Multilingual

An open-source AI model family for on-premises deployment, licensed under Apache 2.0, with language, speech, vision and guardrail models.

273Updated 2 years agoApache-2.0

Linux#Guardrails#Hugging Face integration#LM Studio integration

Open-source reasoning LLM for self-hosted use through vLLM or Transformers, with million-token context and function calling under Apache 2.0.

3.2KUpdated 1 year agoApache-2.0

#Hugging Face integration#Multilingual#Tool calling

An open-source LLM family with reasoning and chat models, Transformers support, and a long-context variant that accepts up to one million tokens.

7.3KUpdated 1 year agoApache-2.0

#Hugging Face integration#Tool calling#Web search

A self-hosted LLM inference server under Apache 2.0, with Docker deployment, multi-GPU support and an OpenAI-compatible chat API. The project is archived.

10.9KUpdated 6 months agoApache-2.0

Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

An open-source ChatGPT alternative for Chrome that uses Ollama models locally to chat across tabs, analyze files, and assist with writing.

1.1KUpdated 7 months agoAGPL-3.0

Web · Browser Extension#Multilingual#Multimodal input#Ollama integration

An open-source on-device AI framework for Android, iOS, desktop and web, with model conversion from PyTorch, TensorFlow and JAX.

3.5KUpdated 20 hours agoApache-2.0

macOS · Windows · Linux · iOS · Android · Web#Agent Skills#Hugging Face integration#Multimodal input

Local AI voice assistant for Linux and Windows with camera vision, persistent memory and MCP tools. Uses Ollama or cloud APIs. MIT licensed.

5.7KUpdated 2 weeks agoMIT

macOS · Windows · Linux · Docker#Home Assistant integration#MCP#Multi-agent workflows

An on-device AI runtime that runs PyTorch models locally on Android, iOS, desktops and embedded hardware, with CPU, GPU, NPU and DSP acceleration.

5.1KUpdated 22 hours ago

macOS · Windows · Linux · iOS · Android · Web#MLX#Multimodal input#OpenAI-compatible API

Self-hosted AI chat interface under Apache 2.0. Connect it to Ollama, llama.cpp or cloud APIs, with optional model routing and MCP tools.

11KUpdated 1 day agoApache-2.0

Docker · Web#llama.cpp backend#MCP#Multi-user access

Favicon of CopilotKit

CopilotKit

2 videos
Self-hostable AI agent SDK with an MIT-licensed core, generative UI, shared app state, and integrations for React, Angular, Vue, Slack, and Teams.

37.6KUpdated 20 hours agoMIT

iOS · Android · Web#Human approval#MCP#Persistent memory

A self-hosted AI gateway under Apache 2.0 that runs in Docker, manages LLM APIs and MCP servers, and supports Kubernetes inference routing.

9.5KUpdated 1 day agoApache-2.0

Docker · Web#Guardrails#MCP#Tool calling

A desktop AI chat app that stores data locally, connects to your own model providers, and supports files, web search and MCP tools.

chatwise.appChat With Your Documents

#MCP#Multimodal input#Tool calling

A self-hosted MCP gateway that combines tool servers behind shared endpoints. Runs in Docker under the MIT license and connects to Cursor and Open WebUI.

2.7KUpdated 3 months agoMIT

Docker · Web#MCP#Multi-user access#Single sign-on