Tools tagged with "Hugging Face integration"

200+ tools
Favicon of Handy

Handy

7 videos
Free, open source desktop dictation for Windows, macOS and Linux. It uses local Whisper or Parakeet models and keeps your voice off the cloud.

32.5KUpdated 3 days agoMIT

macOS · Windows · Linux#GGUF#Hugging Face integration#Voice activity detection

A browser-based local LLM tool that pools laptop, desktop and phone GPUs for chat and coding. Open source under MIT, with no account required.

544Updated 14 hours agoMIT

macOS · iOS · Web#Code execution#Distributed execution#Hugging Face integration

Favicon of exo

exo

1 video
An open-source local LLM runner for macOS and Linux that splits models across devices and works offline with downloaded models. Apache 2.0 licensed.

47.7KUpdated 1 month agoApache-2.0

macOS · Linux · Web#Distributed execution#Hugging Face integration#MLX

Favicon of Jan

Jan

3 videos
A free desktop AI chat app for Windows, macOS and Linux that runs models locally or connects to GPT, Claude and other cloud models.

44.7KUpdated 5 hours ago

macOS · Windows · Linux#Hugging Face integration#MCP#OpenAI-compatible API

A Python library for running diffusion models locally with PyTorch, including Stable Diffusion, LoRA adapters and Apple Silicon support. Apache 2.0 licensed.

34.6KUpdated 1 day agoApache-2.0

macOS#ControlNet#Hugging Face integration#Image-to-image

Favicon of llama.cpp

llama.cpp

11 videos
An open source local LLM engine for GGUF models, with CPU and GPU support, a built-in web UI, and an OpenAI-compatible server.

130KUpdated 1 hour agoMIT

Web#Code execution#GGUF#Hugging Face integration

Favicon of IndexTTS

IndexTTS

1 video
Local text-to-speech software clones voices from one audio clip, supports five languages, and provides separate controls for emotion and speaking speed.

24.2KUpdated 1 day ago

Windows · Linux · Web#Hugging Face integration#Multilingual#Multimodal input

An open-source speech-to-text engine that runs Whisper models locally on desktop and mobile, with CPU-only inference and GPU acceleration. MIT licensed.

54KUpdated 2 days agoMIT

macOS · Windows · Linux · iOS · Android · Docker#Hugging Face integration#Quantization#Streaming inference

Favicon of Wan2.2

Wan2.2

3 videos
Open source video generation models that run on your GPU, with text, image, speech and character animation options under Apache 2.0.

17.7KUpdated 1 week agoApache-2.0

#Hugging Face integration

Favicon of Qwen3

Qwen3

1 video
A language model family with public weights for local CPU or GPU use through Ollama, llama.cpp, and LM Studio, plus server deployment.

27.7KUpdated 9 months ago

#Batch processing#GGUF#Hugging Face integration

Favicon of VibeVoice

VibeVoice

2 videos
Open-source voice AI models for local transcription and speech generation, with MIT licensing, CPU inference, and streaming audio support.

54.5KUpdated 4 weeks agoMIT

#Hugging Face integration#Multilingual#Quantization

An open-source Python library for running and training text, vision, audio and multimodal models locally, with Apache 2.0 licensing and PyTorch support.

166.9KUpdated 3 hours agoApache-2.0

#Hugging Face integration#Multimodal input

Favicon of Hunyuan3D

Hunyuan3D

2 videos
Local AI 3D generator for macOS, Windows and Linux. Create textured meshes from images or text, texture existing models, and work in Blender.

15KUpdated 11 months ago

macOS · Windows · Linux · Web#Hugging Face integration

An on-device document search engine that combines keyword and semantic search with local GGUF models, plus MCP access for AI agents. MIT licensed.

30.1KUpdated 3 weeks agoMIT

macOS#GGUF#Hugging Face integration#Hybrid search

Local AI chat app for macOS that runs GGUF models offline, searches files on-device, and also connects to Claude, ChatGPT, and OpenAI-compatible endpoints.

recurse.chatChat With Your Documents

macOS#GGUF#Hugging Face integration#OpenAI-compatible API

A proprietary local AI annotation tool for training and evaluation data. Works offline, integrates spaCy and Hugging Face, and supports custom Python workflows.

prodi.gyData Labeling and Annotation

Web#Hugging Face integration#Works offline

Self-hosted LLM inference for Kubernetes with NVIDIA, AMD and Apple Silicon support, OpenAI-compatible APIs, and an Apache 2.0 license.

223Updated 17 hours agoApache-2.0

macOS · Linux#GGUF#Git integration#Guardrails

Open-source Python library for synthetic data generation and LLM training, with local models, API-based models, caching, and resumable workflows.

1.1KUpdated 2 years agoMIT

#Hugging Face integration#LoRA#Quantization

Local AI voice conversion software for Windows, Linux and Apple Silicon Macs. Converts speech and singing from a short voice sample; GPL-3.0 and archived.

3.9KUpdated 1 year agoGPL-3.0

macOS · Windows · Linux · Web#Hugging Face integration#Streaming inference#Voice conversion

An open-source singing voice conversion framework that runs fully offline with user-trained models. Licensed under AGPL-3.0; archived and no longer maintained.

28.1KUpdated 3 years agoAGPL-3.0

#Hugging Face integration#ONNX#Voice conversion

Text embedding models for self-hosted multilingual and code search, with custom vector sizes and support for CPU or NVIDIA GPU inference.

2KUpdated 1 year ago

Docker#Batch processing#Hugging Face integration#Multilingual

Local FLUX LoRA training UI for Windows, Linux and Docker, with support for GPUs with 12GB VRAM. Open source under the MIT license.

3.3KUpdated 2 months agoMIT

Windows · Linux · Docker · Web#Hugging Face integration#LoRA

Local AI music generation library built on Stable Diffusion. Run it on your own hardware with CUDA, Apple Silicon or CPU. MIT licensed and no longer maintained.

3.9KUpdated 2 years agoMIT

macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#Multimodal input

Self-hosted RAG chatbot with Ollama and cloud model support, licensed BSD-3-Clause. The project is archived and no longer maintained.

7.7KUpdated 4 months agoBSD-3-Clause

Windows · Docker · Web#Hugging Face integration#Hybrid search#Ollama integration

A local audio transcription CLI that runs Whisper and Distil-Whisper on NVIDIA GPUs or Apple Silicon Macs. Open source under Apache 2.0.

13.1KUpdated 2 years agoApache-2.0

macOS · Windows#Batch processing#Hugging Face integration#Multilingual

An image-to-3D AI model you can run locally on a GPU or CPU, with GLB export, ComfyUI support and hosted access through Stability AI.

1.8KUpdated 2 years ago

macOS · Windows · Linux#Batch processing#Hugging Face integration

A local AI audio generator that turns text into sound effects, music and speech. Runs on CPU, NVIDIA CUDA or Apple Silicon with Hugging Face Diffusers support.

2.6KUpdated 2 years ago

macOS · Linux · Web#Batch processing#Hugging Face integration

A local text embedding model for semantic search and RAG, with adjustable vector sizes, Apache 2.0 licensing, and support for Sentence Transformers.

1.9KUpdated 11 months ago

Docker#Batch processing#Hugging Face integration#ONNX

An open source LLM programming language that combines Python logic with output constraints and supports local llama.cpp and Transformers models or cloud APIs.

4.2KUpdated 1 year agoApache-2.0

Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration

An open-source AI music generator that runs locally on macOS, Windows and Linux, with text or audio style prompts and Apache 2.0 code and DiT weights.

2.3KUpdated 10 months agoApache-2.0

macOS · Windows · Linux · Docker#Hugging Face integration#Multimodal input

An open-source local LLM training toolkit in Python for training GPTs from scratch or fine-tuning GPT-2, with CPU, NVIDIA and Apple Silicon support.

63.5KUpdated 11 months agoMIT

macOS · Windows#Hugging Face integration

A local text-to-audio model for sound effects and music experiments, with CPU or CUDA support and access under the Stability AI Community License.

3.9KUpdated 4 months agoMIT

#Batch processing#Hugging Face integration

A multilingual text embedding model that runs offline on phones, laptops and tablets, with open weights and a quantized memory footprint under 200MB.

5.8KUpdated 11 hours agoApache-2.0

#Hugging Face integration#Multilingual#Quantization

Text embedding models for semantic search, licensed under Apache 2.0, with compact, large and long-context variants for document retrieval.

91Updated 2 years agoApache-2.0

#Hugging Face integration

Alibaba’s GTE models turn text into vectors for retrieval and similarity matching, with downloadable weights and local Python inference.

huggingface.coEmbedding and Reranker Models

#Batch processing#Hugging Face integration#Multilingual

Mixedbread’s mxbai-embed models produce text vectors locally for document retrieval, with Apache-2.0 weights and adjustable embedding dimensions.

mixedbread.comEmbedding and Reranker Models

#Hugging Face integration