9.7KUpdated 1 day ago
macOS · Windows · Linux · Docker · Web#Batch processing#ControlNet#GGUF
Wan2GP brings video, image, music and speech generation to your own computer, with particular attention to GPUs with limited memory. It's for creators who want several media models in one browser interface. The project builds on Wan-Video/Wan2.1.
13.2KUpdated 2 days agoApache-2.0
#ControlNet#Image-to-image#Inpainting
DiffSynth-Studio is a Python diffusion model engine for developers and researchers who want to generate media and train models on their own hardware. It supports large models on consumer GPUs through memory offloading and quantization, with inference and training in the same framework. It's open source under Apache 2.0.
34.6KUpdated 1 day agoApache-2.0
macOS#ControlNet#Hugging Face integration#Image-to-image
Diffusers is an open-source Python library for developers and researchers who want to run diffusion models on their own hardware or build generation features into an application. It uses PyTorch and supports image, video and audio generation. The library is licensed under Apache 2.0 and supports Apple Silicon.
7.3KUpdated 1 week agoApache-2.0
macOS · Windows · Linux · Docker · Web#ControlNet#Image-to-image#Inpainting
SD.Next is a self-hosted web interface for artists, researchers and people who want to generate and edit images or videos on their own hardware. It builds on Automatic1111 WebUI's original codebase and supports Stable Diffusion alongside other diffusion models. It's open source under Apache 2.0.
30.1KUpdated 1 day ago
macOS · Windows#Image-to-image
FaceFusion is an AI face manipulation tool for photos and videos that processes media on your own machine. It's aimed at content creators, VFX artists, and film studios that want control over their footage and personal data. Local processing keeps that media on your hardware rather than sending it to a cloud service.
77KUpdated 22 hours agoApache-2.0
macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image
Unsloth brings model training and everyday AI use into a desktop app for people who want to run models on their own hardware. Its no-code interface covers chat, fine-tuning and media generation on macOS, Windows and Linux. The Unsloth software is open source under Apache 2.0.
1.2KUpdated 9 months agoMIT
macOS · Windows · Linux#Batch processing#Image-to-image
AI Render is a Blender add-on that uses your scene and a text prompt to generate an image with Stable Diffusion. It's for Blender artists who want to use their 3D work as the basis for AI images and explore different treatments within the application they already use.
982Updated 1 year ago
macOS · Windows · Linux · Docker#Home Assistant integration#Image-to-image#Multimodal input
CodeProject.AI Server gives developers a shared API for AI tasks that run on their own hardware. It's a self-hosted service for adding image analysis, text processing and generation to applications. Processing stays on the machine running the server, without cloud calls or sending data outside your device or network.
53.2KUpdated 1 year agoGPL-3.0
macOS · Windows · Linux · Docker · Web#ControlNet#Image-to-image#Inpainting
Fooocus is a free, open-source AI image generator for people who want to create images on their own computer without spending much time tuning settings. It uses Stable Diffusion XL and automatically expands prompts with a local GPT-2 engine. Generation works offline once the required models are downloaded, so prompts and images can stay on your machine.
7.3KUpdated 3 years agoMIT
macOS · Windows#Image-to-image#Inpainting
Auto-Photoshop-StableDiffusion Plugin brings Stable Diffusion generation into Photoshop for artists who want to use AI alongside their existing editing tools. It connects to AUTOMATIC1111 or ComfyUI, so you can generate images within Photoshop and continue editing and saving them there.
3.9KUpdated 2 years agoMIT
macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#Multimodal input
Riffusion is a Python library for generating music and audio on your own hardware using Stable Diffusion. It's for developers and musicians who want to experiment with text-driven sound generation or build it into an app. The hobby project is no longer actively maintained.
966Updated 9 months agoGPL-3.0
Windows#Batch processing#Image-to-image#Inpainting
NMKD Stable Diffusion GUI is a local AI image generator for people who want to create and edit images on a Windows PC. It combines Stable Diffusion generation with inpainting, LoRA training and image post-processing in a desktop interface. It's open source under GPL-3.0.
4.5KUpdated 2 years agoApache-2.0
Docker · Web#Image-to-image
InstantMesh generates a 3D mesh from a single image on your own machine. It's for creators exploring image-based 3D assets and researchers working on 3D reconstruction. The Python project is open source under Apache 2.0 and uses PyTorch with CUDA for GPU processing.
27.3KUpdated 9 months agoMIT
Web#Hugging Face integration#Image-to-image
Stable Diffusion XL is an image generation model for creators and developers who want to generate images on their own hardware or servers. It supports text-to-image generation and image-to-image sampling, so you can start with a written prompt or an existing image.
8.4KUpdated 8 months agoApache-2.0
Web#Image-to-image#LoRA#Multimodal input
Qwen-Image is an open-source image generation and editing model you can deploy locally. It's for developers and creators who want to generate images from text or revise existing pictures on their own hardware. Its text rendering capabilities, especially for Chinese, make it relevant for images that need readable lettering alongside visual content.
73.5KUpdated 4 years ago
#Guardrails#Hugging Face integration#Image-to-image
Stable Diffusion 1.5 is an AI image generation model for creators and developers who want to generate images on their own hardware. It turns text prompts into images and supports text-guided changes to existing pictures, including turning rough sketches into detailed artwork. Local inference keeps that image generation work on your machine.
16.2KUpdated 1 day agoApache-2.0
Windows · iOS · Android#Image-to-image#Multimodal input#ONNX
MNN is a lightweight C++ framework for developers who want AI models to run on phones, PCs and embedded devices. It handles inference and training on the device, with a focus on small application footprints and hardware acceleration. The project is open source under Apache 2.0, and Alibaba uses it in apps including Taobao, Youku and DingTalk.
4.6KUpdated 2 years agoApache-2.0
Web#ControlNet#Hugging Face integration#Image-to-image
Kolors is a text-to-image model for people who want to generate photorealistic images on their own hardware, including work with Chinese prompts and Chinese cultural content. Developed by Kuaishou, it understands prompts in Chinese and English and can render text in both languages within generated images.
23.3KUpdated 1 year agoApache-2.0
macOS · Windows · Web#Batch processing#Image-to-image#Inpainting
IOPaint is a free, self-hosted AI image editor for people who want to remove objects, replace parts of a picture or extend it beyond its original edges on their own hardware. The project is archived and no longer maintained. It's open source under Apache 2.0, with a browser interface and support for CPU, GPU and Apple Silicon hardware.
4.3KUpdated 10 months agoMIT
Web#Hugging Face integration#Image-to-image#LoRA
OmniGen is a local AI image generation model that handles text prompts, reference images, and image editing within one model. It's for creators who want to reuse subjects across images and developers building image tools on their own hardware. The code is open source under the MIT license.
8KUpdated 5 days agoGPL-3.0
macOS#ControlNet#Image-to-image#Multimodal input
Mochi Diffusion is a native macOS image generator for people who want to create images on an Apple Silicon Mac without sending prompts or results to a cloud service. It runs Stable Diffusion and FLUX.2 Klein locally, with completely offline generation. The app collects no telemetry or other data and is open source under GPL-3.0.
2.9KUpdated 20 hours agoAGPL-3.0
macOS · Docker · Web#ControlNet#Distributed execution#Human approval
SimpleTuner is an open-source toolkit for fine-tuning image, video and audio generation models on your own hardware or GPU servers. It's for creators and researchers adapting models to their datasets, and teams sharing training infrastructure. A web dashboard manages training jobs.
9.2KUpdated 7 days agoApache-2.0
Docker#Image-to-image#Inpainting#Multimodal input
ModelScope combines a hosted model and dataset hub with a Python library you can run locally. It's for developers and researchers who want to use AI models in their own applications, fine-tune them on their own data, or compare their performance. The library is open source under Apache 2.0.
29.9KUpdated 3 weeks ago
macOS · Linux · iOS · Android · Web#Image-to-image#ONNX#Quantization
InsightFace is a face analysis toolkit for developers and teams building identity verification, access control, or face editing software. The code uses the MIT license. Its Python tools and self-hosted recognition server run inference on your own hardware. It also offers commercial models and API access for face swapping and deepfake detection.
10.5KUpdated 3 weeks ago
macOS · Windows · Linux · Web#Batch processing#ControlNet#Image-to-image
Easy Diffusion runs Stable Diffusion on your own computer through a browser interface. It's for people who want to generate and edit images locally without assembling the software components themselves. The free distribution bundles the required software and works on Windows, Linux and macOS.
18.2KUpdated 10 months ago
#Image-to-image#Inpainting
CodeFormer restores degraded faces in photos and videos using a learned library of facial details and a Transformer model. It's for people restoring images on their own hardware, developers adding face enhancement to a project, and researchers working on image restoration. Its main distinction is an adjustable balance between visual quality and fidelity to the input face.
13KUpdated 1 year agoAGPL-3.0
Windows · Web#ControlNet#GGUF#Image-to-image
Stable Diffusion WebUI Forge runs image generation on your own hardware through a browser interface. It builds on Stable Diffusion WebUI and suits people who want its image creation tools with more control over GPU memory use, as well as developers extending those tools. It's open source under AGPL-3.0.
165.2KUpdated 2 years agoAGPL-3.0
macOS · Windows · Linux · Web#Batch processing#Code execution#Image-to-image
Stable Diffusion web UI (AUTOMATIC1111) is a browser interface for generating and editing images with models running on your own hardware. It's for artists and anyone who wants control over prompts, models and image variations. The software is open source under AGPL-3.0.
1.5KUpdated 2 years agoMIT
#ControlNet#Hugging Face integration#Image-to-image
Stable Diffusion 3.5 is a family of text-to-image models for people building image tools or producing visual work on their own infrastructure. It generates photography, paintings, line art and 3D-style images from prompts, with an emphasis on following the requested subject and composition.
12.2KUpdated 3 days agoMIT
macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#LoRA
AI Toolkit (ostris) is an MIT-licensed training suite for people who want to fine-tune image and video models on their own hardware or a self-hosted server. It targets consumer NVIDIA GPUs and runs on Linux and Windows, including ARM64 Linux systems such as DGX Spark. An experimental installer also supports Apple Silicon Macs. GPU memory needs depend on the model and training task.
16.3KUpdated 1 week agoApache-2.0
Web#Hugging Face integration#Image-to-image#Multilingual
Transformers.js is a JavaScript library for developers building web apps that run AI models on the user's device. Inference happens in the browser, so an app doesn't need a separate model server to process its inputs. The library is open source under Apache 2.0.
575Updated 1 day agoGPL-3.0
macOS · Linux · iOS · Docker#Image-to-image#Inpainting#LoRA
Draw Things is an AI image generation app for iPhone, iPad and Mac that keeps generation on your device and works offline. It's for people who want to create and edit images without sending that work to a cloud service, including artists developing character concepts or trying out apparel designs.
13.6KUpdated 3 years agoAGPL-3.0
macOS#ControlNet#Image-to-image#Inpainting
DiffusionBee is an open-source AI art app for Mac users who want to generate and edit images on their own computer. It runs Stable Diffusion offline, with image generation processed on the device. Model downloads require network access, and optional image uploads can send images externally. Its visual interface suits artists and designers who want local image tools without working through code.
8.2KUpdated 2 weeks agoGPL-3.0
macOS#Image-to-image#Inpainting#Visual workflows
Dream Textures is an open-source Blender add-on for artists who want to generate images and apply AI textures to 3D work inside Blender. It runs Stable Diffusion on your own machine and connects image generation to scene depth, texture projection, and animation rendering.
5.8KUpdated 2 days agoMIT
macOS · Windows · Linux · Docker · Web#GGUF#Image-to-image#llama.cpp backend
llama-swap is a self-hosted proxy for people running several AI models on their own hardware. It starts the model server a request needs and swaps out another when necessary, so you don't have to keep every model loaded or manage separate API connections in your apps.
37.7KUpdated 2 years ago
Linux#Batch processing#Image-to-image
GFPGAN is an open-source AI face restoration tool for people repairing poor-quality photos and developers adding restoration to image workflows. It runs locally with Python and PyTorch, with Linux support and optional NVIDIA GPU acceleration through CUDA. Its face models use knowledge learned by a pretrained generative model such as StyleGAN2 to reconstruct facial detail.