Tools tagged with "Hugging Face integration"

200+ tools
Open-source Python toolkit for local AI audio generation, fine-tuning and training, with a Gradio interface and support for Stable Audio Open.

3.9KUpdated 4 months agoMIT

Web#Hugging Face integration

Open-source AI framework for semantic search, RAG and agents. Runs locally or in Docker, with Hugging Face, llama.cpp and cloud models via LiteLLM.

13KUpdated 23 hours agoApache-2.0

Docker#Agent Skills#Hugging Face integration#Knowledge graphs

Favicon of YuE

YuE

1 video
An open-source AI music generator that turns lyrics into songs. Run it locally on Linux with an NVIDIA GPU, or use the hosted demo.

10.6KUpdated 2 days agoApache-2.0

Linux · Web#Hugging Face integration#Multilingual#Multimodal input

A local LLM family for developers and researchers, with downloadable weights, text and vision models, and custom licensing for research and commercial use.

7.7KUpdated 12 months ago

#Hugging Face integration#Multimodal input#Quantization

Open-source LLM serving toolkit for your own GPU servers, with quantization, text and vision models, and OpenAI-compatible APIs. Apache 2.0 licensed.

8.1KUpdated 3 days agoApache-2.0

#Batch processing#Distributed execution#Hugging Face integration

DeepSeek-V3 and R1 are downloadable language-model families for local text generation and reasoning.

91.9KUpdated 1 year agoMIT

#Hugging Face integration

Open-source vision-language models for visual chat, document questions and image retrieval, with downloadable weights and Hugging Face Transformers support.

10.2KUpdated 1 year agoMIT

#Hugging Face integration#Multimodal input

LLM training framework with ready-made research scripts, NVIDIA GPU parallelism, and Hugging Face checkpoint conversion through Megatron Bridge.

18KUpdated 1 day ago

Docker#Distributed execution#Hugging Face integration#Quantization

Favicon of smolagents

smolagents

1 video
An open-source Python AI agent library with local models through Transformers or Ollama, cloud API support, and Docker sandboxing.

29.6KUpdated 1 week agoApache-2.0

#Code execution#Hugging Face integration#MCP

An open-source LLM evaluation toolkit for macOS and Linux. Test local models on CPU or GPUs, or evaluate hosted APIs, under the MIT license.

2.5KUpdated 2 days agoMIT

macOS · Linux#Hugging Face integration#Multilingual

Favicon of MLX LM

MLX LM

2 videos
A Python package for local LLM inference and fine-tuning on Apple Silicon, built on MLX with Hugging Face model support and an MIT license.

7.2KUpdated 1 day agoMIT

macOS#Batch processing#Distributed execution#Hugging Face integration

An open source 3B local LLM that runs on CPU or CUDA GPU and supports direct or reasoning responses, six languages and a 128k context window.

3.9KUpdated 1 week agoApache-2.0

#Hugging Face integration#Multilingual#Tool calling

Self-hosted text-to-speech server runs Chatterbox models on CPU or GPU, with voice cloning, audiobook generation and an OpenAI-compatible API.

1.5KUpdated 4 months agoMIT

macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#OpenAI-compatible API

Favicon of PEFT

PEFT

1 video
An open-source Python library for adapting LLMs and diffusion models with less memory and storage, under Apache 2.0.

21.7KUpdated 1 day agoApache-2.0

macOS#Distributed execution#Hugging Face integration#LoRA

An open-source segmentation model for images and videos. Run it on your own GPU machine, track objects across frames, and refine masks with prompts.

19.9KUpdated 2 years agoApache-2.0

Web#Hugging Face integration

Favicon of Axolotl

Axolotl

1 video
Apache 2.0 LLM fine-tuning framework for local or cloud GPUs, supporting NVIDIA, AMD, LoRA, QLoRA and multimodal training.

12.5KUpdated 2 days agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

Open-source speaker diarization toolkit with local PyTorch models, CUDA GPU support, and an optional hosted service that processes audio on pyannoteAI servers.

10.6KUpdated 3 months agoMIT

#Hugging Face integration#Speaker diarization#Voice activity detection

A self-hosted OCR model that extracts text and page structure from PDFs and images, with vLLM, Hugging Face Transformers and CPU inference support.

9.2KUpdated 6 months agoMIT

Docker#Hugging Face integration#Multilingual#Multimodal input

Favicon of OpenVINO

OpenVINO

2 videos
Open-source AI inference toolkit for local or self-hosted deployment on Linux, Windows and macOS, with CPU, Intel GPU and NPU support.

10.9KUpdated 1 day agoApache-2.0

macOS · Windows · Linux#Hugging Face integration#Multimodal input#ONNX

Open-source AI podcast generator in Python with local HuggingFace models for transcripts and cloud speech services for multilingual audio.

6.6KUpdated 5 months agoApache-2.0

Docker · Web#Hugging Face integration#Multilingual#Multimodal input

Open-source Python library for compressing local or Hugging Face LLM checkpoints, with Apache 2.0 licensing and vLLM-compatible output.

3.8KUpdated 1 day agoApache-2.0

#Hugging Face integration#Quantization

Favicon of KServe

KServe

2 videos
Self-hosted AI model serving platform for Kubernetes. Serve LLMs and predictive models with vLLM, Hugging Face support and an OpenAI-compatible API.

6KUpdated 23 hours agoApache-2.0

#Hugging Face integration#ONNX#OpenAI-compatible API

A self-hosted AI video generation model with ComfyUI and Diffusers support, image animation, video editing, and fine-tuning tools.

11KUpdated 9 months agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input

An open-source Python AI framework for self-hosted agents and document search, with local model support and an Apache 2.0 license.

26.6KUpdated 1 day agoApache-2.0

Docker#Guardrails#Hugging Face integration#Hybrid search

Python library for running and training embedding and reranker models locally, with Apache 2.0 licensing and pretrained models on Hugging Face.

19.1KUpdated 1 week agoApache-2.0

#Hugging Face integration#Multilingual#Multimodal input

Open-source Rust ML framework for running models locally on CPUs, NVIDIA GPUs or in browsers, with Apache 2.0 licensing and quantized LLM support.

21.1KUpdated 2 days agoApache-2.0

macOS · Web#GGUF#Hugging Face integration#Multilingual

A local LLM merging toolkit that combines model weights without extra training. Runs on CPU or GPU and supports Llama, Mistral and PyTorch checkpoints.

7.4KUpdated 3 weeks agoLGPL-3.0

#Hugging Face integration#LoRA

Self-hosted embedding and reranking API with MIT licensing, Hugging Face models, and CPU, NVIDIA, AMD and Apple MPS support.

2.9KUpdated 6 months agoMIT

macOS · Docker#Batch processing#Hugging Face integration#Multimodal input

Run AI models through Docker Desktop, Docker Engine or a standalone binary, with local inference and OpenAI and Ollama compatible APIs.

655Updated 2 days agoApache-2.0

macOS · Windows · Linux#GGUF#Hugging Face integration#llama.cpp backend

An open-source Python library for local speech transcription with Whisper models. It runs on CPUs or NVIDIA GPUs and uses CTranslate2.

25.6KUpdated 3 hours agoMIT

#Batch processing#Hugging Face integration#Quantization

Open-weight local LLMs under Apache 2.0, with 20B and 120B models that work with Ollama, LM Studio and vLLM.

20.4KUpdated 2 months agoApache-2.0

macOS · Linux#Hugging Face integration#LM Studio integration#Ollama integration

Local AI app for Android and iOS. Run Gemma 4 on-device, ask questions about photos, transcribe audio and compare model performance.

24.8KUpdated 22 hours agoApache-2.0

iOS · Android#Agent Skills#Hugging Face integration#Multilingual

Local AI image manager for photo culling and dataset curation. Run it as a desktop app or self-hosted server, with on-device models and ComfyUI integration.

89Updated 23 hours agoGPL-3.0

macOS · Windows · Linux · Docker · Web#Batch processing#Hugging Face integration#ONNX