Tools tagged with "Batch processing"

100+ tools
Self-hosted LLM inference engine for Hugging Face models, with OpenAI-compatible APIs, multimodal support, and CPU or GPU execution under AGPL-3.0.

1.9KUpdated 3 weeks agoAGPL-3.0

macOS · Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

A local vision-language model for image analysis, document extraction and video understanding, with Apache 2.0 weights and Hugging Face Transformers support.

huggingface.coOCR and Document Scanning

Linux#Batch processing#Hugging Face integration#Multimodal input

An open source LLM evaluation framework for local models and hosted APIs, with GGUF, Hugging Face transformers and llama.cpp support. MIT licensed.

14.1KUpdated 2 weeks agoMIT

macOS#Batch processing#GGUF#Hugging Face integration

An open-source vector search library for local apps and servers, with custom metrics, disk-backed indexes, and Apache 2.0 licensing.

4.3KUpdated 4 weeks agoApache-2.0

macOS · Windows · Linux · iOS · Android#Batch processing#Semantic search

An open-source image and text model for local image classification without task-specific training. Runs through PyTorch on CPU or CUDA GPUs under MIT.

34.4KUpdated 6 months agoMIT

#Batch processing#Multimodal input

A free, self-hosted AI image editor for erasing objects and extending images. Runs on CPU, GPU or Apple Silicon. Apache 2.0; archived and unmaintained.

23.3KUpdated 1 year agoApache-2.0

macOS · Windows · Web#Batch processing#Image-to-image#Inpainting

A local LLM inference library for Windows and Linux with NVIDIA GPUs, MIT licensing, GPTQ and EXL2 support, and an OpenAI-compatible API through TabbyAPI.

4.6KUpdated 7 months agoMIT

Windows · Linux#Batch processing#Quantization#Speculative decoding

A self-hosted LLM inference server under Apache 2.0, with Docker deployment, multi-GPU support and an OpenAI-compatible chat API. The project is archived.

10.9KUpdated 6 months agoApache-2.0

Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

A local multimodal AI model for answering image questions and generating pictures, with downloadable weights and a Gradio interface.

17.8KUpdated 2 years agoMIT

Web#Batch processing#Hugging Face integration#Multimodal input

An open-source image generation framework for local use, with ComfyUI, Diffusers and fine-tuning support. Code uses the Apache 2.0 license.

1KUpdated 4 months agoApache-2.0

Linux · Web#Batch processing#Hugging Face integration#LoRA

An open-source wake word library for local voice apps, with English models, custom phrase training, and ONNX support on Linux and Windows.

2.8KUpdated 9 months agoApache-2.0

Windows · Linux#Batch processing#ONNX#Voice activity detection

A free, open-source meme search engine that runs locally in Docker, with AI image descriptions, semantic search and optional OpenAI-compatible vision APIs.

752Updated 3 weeks agoApache-2.0

Linux · Docker · Web · Browser Extension#Batch processing#OpenAI-compatible API#Quantization

A desktop image dataset editor with local AI captioning on CPU or NVIDIA GPU. Runs on Windows, Linux and macOS under GPL-3.0.

1.4KUpdated 12 months agoGPL-3.0

macOS · Windows · Linux#Batch processing#Multimodal input

An open-source document extraction engine with a Rust core, CPU-only processing, Docker deployment, and support for Ollama, LM Studio, and vLLM.

9.4KUpdated 2 days agoMIT

macOS · Windows · Linux · Android · Docker · Web#Batch processing#LM Studio integration#MCP

Open-source speech recognition model for Mandarin, Cantonese, English, Japanese and Korean. Runs locally on CPU or GPU under the MIT license.

9.4KUpdated 3 weeks agoMIT

Docker#Batch processing#GGUF#Hugging Face integration

A self-hosted AI add-on for paperless-ngx that extracts text and organizes documents using local Ollama models or cloud APIs. Open source under MIT.

2.7KUpdated 6 days agoMIT

Docker · Web#Batch processing#Human approval#Multimodal input

Local AI transcription app for macOS that keeps audio on your device, supports Whisper, Qwen and Parakeet, and offers optional cloud services.

goodsnooze.gumroad.comDictation and Voice Typing

macOS · iOS#Batch processing#Multilingual#Ollama integration

Open-source LLM serving infrastructure for Kubernetes with multi-node inference, demand-based autoscaling, LoRA management and vLLM integration.

5.1KUpdated 21 hours agoApache-2.0

#Batch processing#Distributed execution#LoRA

An open-source Python library for synthetic data and structured extraction, with Ollama, vLLM, cloud APIs and managed fine-tuning integrations.

1.7KUpdated 2 days agoApache-2.0

#Batch processing#Code execution#Multimodal input

An open-source video translation app for Windows, macOS and Linux, with local Qwen3 speech recognition, bilingual subtitles and optional dubbing.

18.5KUpdated 3 days agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#MLX#Multilingual

A self-hosted model serving library for Python and LLM APIs. Run it on a laptop or cluster with request batching, streaming, and CPU or GPU resources.

44KUpdated 19 hours agoApache-2.0

#Batch processing#Hugging Face integration#ONNX

Self-hosted AI agent platform for visual workflows, research and development. Runs on macOS, Windows, Linux or Docker under Apache 2.0.

34.4KUpdated 2 months agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#MCP

Self-hosted AI compute orchestration under MPL-2.0 for training and inference on GPU clouds, Kubernetes, VMs and bare-metal servers.

2.3KUpdated 1 day agoMPL-2.0

macOS · Windows · Linux · Docker#Agent Skills#Batch processing#Multi-user access

Local AI image search finds photos by text or example image. Runs offline on Linux, Windows and Apple Silicon macOS under the MIT license.

1KUpdated 7 days agoMIT

macOS · Windows · Linux#Batch processing#Multimodal input#Semantic search

A local LLM inference engine for sparse models, with CPU and GPU support on Linux and Windows. Open source under MIT, with CPU-only support on Apple Silicon.

9.8KUpdated 5 months agoMIT

macOS · Windows · Linux#Batch processing#GGUF#Hugging Face integration

Local OCR model that extracts plain or formatted text from images, supports multi-page recognition, and works with Hugging Face Transformers.

8.2KUpdated 2 years ago

#Batch processing#GGUF#Hugging Face integration

An open-source Python tool for AI agent evaluation and tracing, with MIT licensing, OpenTelemetry support and a database you control.

3.6KUpdated 1 day agoMIT

#Batch processing#LLM tracing#MCP

Favicon of Parakeet

Parakeet

1 video
English speech-to-text model that runs locally through NVIDIA NeMo on Linux, with punctuation, word timestamps and CC BY 4.0 licensed weights.

18.5KUpdated 1 day agoApache-2.0

Linux · Docker#Batch processing#Hugging Face integration

Local OCR model that converts images and PDFs to text or Markdown. Runs on NVIDIA GPUs with vLLM or Transformers under the MIT license.

23.9KUpdated 8 months agoMIT

Linux#Batch processing#Hugging Face integration#Multimodal input

Local speech recognition models for English transcription, built on Whisper. Run on CPU or CUDA GPUs with Hugging Face Transformers under the MIT license.

4.1KUpdated 2 years agoMIT

#Batch processing#Hugging Face integration

Local OCR and document parsing software that converts PDFs and images to Markdown, recognizes tables and formulas, and runs on NVIDIA GPUs.

6.7KUpdated 2 months agoApache-2.0

Windows · Docker · Web#Batch processing#Hugging Face integration#Multilingual

An open-source macOS dictation app that runs Whisper and Parakeet models locally on Apple Silicon and transcribes microphone recordings or audio files.

3KUpdated 3 weeks agoMIT

macOS#Batch processing#Hugging Face integration#Multilingual

A self-hosted LLM API server that runs ExLlamaV3 models on your hardware, with OpenAI-compatible endpoints and an AGPL-3.0 license.

1.4KUpdated 2 days agoAGPL-3.0

Windows · Linux · Docker#Batch processing#Distributed execution#Hugging Face integration

Open-source AI data processing framework for local machines and Ray clusters, with multimodal cleaning, deduplication and Apache 2.0 licensing.

7.1KUpdated 2 days agoApache-2.0

Docker#Batch processing#Distributed execution#Multimodal input

Favicon of llm-d

llm-d

1 video
An open-source LLM inference stack for self-hosted Kubernetes clusters, with vLLM and SGLang backends and support for GPUs, TPUs, XPUs and CPUs.

4.7KUpdated 1 day agoApache-2.0

#Batch processing#Distributed execution#OpenAI-compatible API

A free local AI image generator that runs Stable Diffusion on Windows, Linux and macOS, with a browser interface for prompts, image editing and ControlNet.

10.5KUpdated 3 weeks ago

macOS · Windows · Linux · Web#Batch processing#ControlNet#Image-to-image