Tools tagged with "Batch processing"

100+ tools
Favicon of Wan2GP

Wan2GP

5 videos
Local AI media generator with a browser interface, support for NVIDIA and AMD GPUs, and select models that run with 6 GB of VRAM.

9.7KUpdated 1 day ago

macOS · Windows · Linux · Docker · Web#Batch processing#ControlNet#GGUF

Favicon of Firecrawl

Firecrawl

4 videos
An open source web scraping API you can self-host or use as a hosted service to turn websites into Markdown, JSON, HTML, or screenshots.

186.5KUpdated 1 day agoAGPL-3.0

Docker#Batch processing#MCP#Structured output

Favicon of vLLM

vLLM

9 videos
An open source LLM serving engine that runs on your hardware, supports NVIDIA and AMD GPUs, and provides an OpenAI-compatible API.

93KUpdated 2 hours agoApache-2.0

macOS · Docker#Batch processing#Distributed execution#GGUF

Favicon of InvokeAI

InvokeAI

2 videos
A self-hosted AI image generator that runs on Windows, macOS, Linux and Docker, with canvas editing, visual workflows and an Apache 2.0 license.

28.3KUpdated 3 days agoApache-2.0

macOS · Windows · Linux · Docker · Web#Batch processing#ControlNet#GGUF

Self-hosted text-to-speech with voice cloning, multilingual speech and emotion control. Code and weights use the FISH AUDIO RESEARCH LICENSE.

32.9KUpdated 2 weeks ago

#Batch processing#Multilingual#Multimodal input

Favicon of SGLang

SGLang

3 videos
An open-source inference framework for serving language and multimodal models on your own hardware, with an OpenAI-compatible API.

36.7KUpdated 2 hours agoApache-2.0

#Batch processing#Distributed execution#LoRA

Favicon of Qwen3

Qwen3

1 video
A language model family with public weights for local CPU or GPU use through Ollama, llama.cpp, and LM Studio, plus server deployment.

27.7KUpdated 9 months ago

#Batch processing#GGUF#Hugging Face integration

Favicon of rembg

rembg

1 video
Local AI background removal tool with ONNX models, CPU and GPU support, and a Python library or self-hosted HTTP API. MIT licensed.

24.9KUpdated 1 week agoMIT

Windows · Docker · Web#Batch processing#ONNX

A self-hosted photo server with AI tagging that keeps images out of the cloud. Runs on Windows, Linux, macOS or Docker.

picapport.deAI Photo Libraries

macOS · Windows · Linux · Docker · Web#Batch processing#Multi-user access#Role-based access

Open-source dictation app for macOS, Windows, Linux and iOS. Transcribe offline with local models or choose cloud processing with your own API keys.

8.9KUpdated 10 hours agoMIT

macOS · Windows · Linux · iOS#Batch processing#MCP#Multilingual

Local AI upscaler for Windows that enlarges images, animations and video, interpolates frames, and supports AMD, NVIDIA and Intel GPUs.

17.1KUpdated 2 weeks ago

Windows#Batch processing#Works offline

A paid audio transcription app that runs Whisper locally on macOS, iOS and visionOS, with multilingual transcription and subtitle exports.

sindresorhus.comOn-Device and In-Browser AI

macOS · iOS#Batch processing#Multilingual

Self-hosted AI application platform for enterprise workflows, RAG and agents. Runs through Docker with a browser interface under Apache 2.0.

12KUpdated 1 week agoApache-2.0

Docker · Web#Batch processing#Human approval#Multi-agent workflows

Local AI music separation software splits songs into stems on Windows, macOS and Linux. MIT licensed, with CPU or CUDA GPU processing.

10.4KUpdated 3 years agoMIT

macOS · Windows · Linux · Docker#Batch processing#Quantization

A Blender add-on that generates images from scenes and text prompts, with local Stable Diffusion support through Automatic1111 on Windows, macOS and Linux.

1.2KUpdated 9 months agoMIT

macOS · Windows · Linux#Batch processing#Image-to-image

Favicon of digiKam

digiKam

1 video
An open-source photo manager for Linux, Windows and macOS with local AI tagging, face recognition, RAW editing and batch processing.

808Updated 7 years agoGPL-2.0

macOS · Windows · Linux#Batch processing

Local AI PDF parser that converts scientific papers to Markdown with LaTeX math and tables. Runs on CPU or GPU, with MIT code and CC-BY-NC weights.

10.1KUpdated 2 years agoMIT

Windows#Batch processing

Text embedding models for self-hosted multilingual and code search, with custom vector sizes and support for CPU or NVIDIA GPU inference.

2KUpdated 1 year ago

Docker#Batch processing#Hugging Face integration#Multilingual

Local AI image generator for Windows 10/11 with Stable Diffusion, inpainting and LoRA training. Open source under GPL-3.0, with NVIDIA and AMD GPU support.

966Updated 9 months agoGPL-3.0

Windows#Batch processing#Image-to-image#Inpainting

A local audio transcription CLI that runs Whisper and Distil-Whisper on NVIDIA GPUs or Apple Silicon Macs. Open source under Apache 2.0.

13.1KUpdated 2 years agoApache-2.0

macOS · Windows#Batch processing#Hugging Face integration#Multilingual

An image-to-3D AI model you can run locally on a GPU or CPU, with GLB export, ComfyUI support and hosted access through Stability AI.

1.8KUpdated 2 years ago

macOS · Windows · Linux#Batch processing#Hugging Face integration

A local AI audio generator that turns text into sound effects, music and speech. Runs on CPU, NVIDIA CUDA or Apple Silicon with Hugging Face Diffusers support.

2.6KUpdated 2 years ago

macOS · Linux · Web#Batch processing#Hugging Face integration

An open-source image-to-3D model that runs locally with Python and PyTorch. MIT-licensed code and weights, with about 6GB VRAM for default inference.

7KUpdated 4 months agoMIT

#Batch processing

A local text embedding model for semantic search and RAG, with adjustable vector sizes, Apache 2.0 licensing, and support for Sentence Transformers.

1.9KUpdated 11 months ago

Docker#Batch processing#Hugging Face integration#ONNX

An open source LLM programming language that combines Python logic with output constraints and supports local llama.cpp and Transformers models or cloud APIs.

4.2KUpdated 1 year agoApache-2.0

Windows · Linux · Web · VS Code#Batch processing#Guardrails#Hugging Face integration

An open-source Python library for preparing LLM training data locally or on Slurm and Ray clusters, with filtering, deduplication and generation.

3.4KUpdated 1 day agoApache-2.0

#Batch processing#Distributed execution#Multilingual

A local text-to-audio model for sound effects and music experiments, with CPU or CUDA support and access under the Stability AI Community License.

3.9KUpdated 4 months agoMIT

#Batch processing#Hugging Face integration

Alibaba’s GTE models turn text into vectors for retrieval and similarity matching, with downloadable weights and local Python inference.

huggingface.coEmbedding and Reranker Models

#Batch processing#Hugging Face integration#Multilingual

A local document retrieval library that matches text queries to page images without OCR. MIT-licensed Python code supports NVIDIA and Apple Silicon GPUs.

2.8KUpdated 1 month agoMIT

macOS#Batch processing#Hugging Face integration#LoRA

Local LLM quantization library for smaller model weights and inference on your hardware. MIT licensed, with CPU and GPU support; archived and unmaintained.

2.3KUpdated 1 year agoMIT

Linux#Batch processing#GGUF#Hugging Face integration

Higgs Audio V2, now Higgs TTS 2, is a downloadable speech model for expressive narration, multilingual dialogue and voice cloning.

8.4KUpdated 4 months agoApache-2.0

#Batch processing#Hugging Face integration#Multilingual

Self-hosted LLM gateway with evaluation, A/B testing, and Ollama support. Open source under Apache 2.0; archived and no longer maintained.

11.7KUpdated 4 months agoApache-2.0

Docker · Web#Batch processing#LLM tracing#Multimodal input

A self-hosted Python library for PostgreSQL RAG and semantic search, with Ollama and cloud embedding providers. Open source, archived and no longer maintained.

5.8KUpdated 4 months agoPostgreSQL

Docker#Batch processing#Ollama integration#RAG

Self-hosted RAG framework for document Q&A, with Docker, Ollama and Infinity support. Apache 2.0 licensed; archived and no longer maintained.

4.4KUpdated 7 months agoApache-2.0

Docker · Web#Batch processing#Multimodal input#Ollama integration

A local text-to-image model for Chinese and English prompts, with Diffusers support, a community ComfyUI wrapper, and Apache 2.0 code.

1.1KUpdated 2 years agoApache-2.0

Web#Batch processing#Hugging Face integration#LoRA

Self-hosted AI application server with an OpenAI-compatible API, local Ollama and vLLM backends, document search and agent tool calling. MIT licensed.

8.4KUpdated 21 hours agoMIT

#Agent Skills#Batch processing#Guardrails