DiffSynth-Studio is a Python diffusion model engine for developers and researchers who want to generate media and train models on their own hardware. It supports large models on consumer GPUs through memory offloading and quantization, with inference and training in the same framework. It's open source under Apache 2.0.
Its model support covers image generation and editing with Qwen-Image and FLUX, video generation with Wan, and music generation with ACE-Step. MiniMax-H3 supports video with audio, including keyframe guidance and reference-based generation. Other supported tasks include character animation, image layer separation and image quality evaluation.
Memory management moves model weights between disk, system memory and GPU memory to reduce VRAM demand. Quantization reduces memory use for both generation and LoRA training. Supported backends include bitsandbytes, torchao and EntroPack, which can compress weights with either exact recovery or controlled loss.
For training, the framework supports full model updates, LoRAs and adapters with additional inputs. CPU offloading makes large-model LoRA training possible on consumer GPUs, while split training separates data preparation from the computations that update model weights.
DiffSynth-ComfyUI brings the engine's inference capabilities into ComfyUI workflows. DiffSynth-WebUI provides a separate interface for privately deployed LoRA training. ModelScope AIGC Zone and Civision offer hosted experiences powered by the engine.
Some older features are no longer maintained and require historical releases. A small development team limits the pace of new features and issue fixes. Model weights have separate licenses.
Claim this page and we'll verify you by hand. DiffSynth-Studio gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find DiffSynth-Studio?Promote it
Something wrong or outdated on this page?
34.6KUpdated 1 day agoApache-2.0
macOS#ControlNet#Hugging Face integration#Image-to-image
Diffusers is an open-source Python library for developers and researchers who want to run diffusion models on their own hardware or build generation features into an application. It uses PyTorch and supports image, video and audio generation. The library is licensed under Apache 2.0 and supports Apple Silicon.
13.6KUpdated 3 years agoAGPL-3.0
macOS#ControlNet#Image-to-image#Inpainting
575Updated 24 hours agoGPL-3.0
macOS · Linux · iOS · Docker#Image-to-image#Inpainting#LoRA
9.2KUpdated 2 weeks agoApache-2.0
#ControlNet#LoRA#Multimodal input
7.2KUpdated 6 days agoApache-2.0
Windows · Linux#ControlNet#Inpainting#LoRA
9.7KUpdated 1 day ago
macOS · Windows · Linux · Docker · Web#Batch processing#ControlNet#GGUF
DiffusionBee is an open-source AI art app for Mac users who want to generate and edit images on their own computer. It runs Stable Diffusion offline, with image generation processed on the device. Model downloads require network access, and optional image uploads can send images externally. Its visual interface suits artists and designers who want local image tools without working through code.
Draw Things is an AI image generation app for iPhone, iPad and Mac that keeps generation on your device and works offline. It's for people who want to create and edit images without sending that work to a cloud service, including artists developing character concepts or trying out apparel designs.
Sana is an open-source framework for running image and video generation on your own hardware, with image models small enough for laptop GPUs. It's aimed at creators who want local AI generation and developers who need training and inference pipelines for their own models. The code uses the Apache 2.0 license.
sd-scripts is a collection of Python scripts for training and generating images with models on your own hardware. It's aimed at people who want to customize image models through LoRA training or deeper fine-tuning and are comfortable working with scripts. The project is open source under the Apache 2.0 license.
Wan2GP brings video, image, music and speech generation to your own computer, with particular attention to GPUs with limited memory. It's for creators who want several media models in one browser interface. The project builds on Wan-Video/Wan2.1.