
Stable Diffusion 3.5 is a family of text-to-image models for people building image tools or producing visual work on their own infrastructure. It generates photography, paintings, line art and 3D-style images from prompts, with an emphasis on following the requested subject and composition.
The Large model is aimed at professional image work. Large Turbo favors faster generation, while Medium balances image quality with customization. Medium targets consumer hardware. For more directed results, Large can use Blur, Canny and Depth ControlNets to guide images from visual inputs.
You can run the models in your own environment with a self-hosted license. Stability AI also provides remote access through its API, cloud partners and the web-based Stable Assistant. Its image editing APIs cover tasks such as removing objects, extending an image and changing backgrounds; those services are separate from running the image models yourself.
The Python reference implementation is MIT-licensed and contains code for basic inference, including the text encoders, image decoder and diffusion model. Model weights are obtained separately under a model license. The reference code is intended for organizations implementing the models; ComfyUI is an alternative for running them.
Claim this page with an email at stability.ai. Stable Diffusion 3.5 gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Stable Diffusion 3.5?Promote it
Something wrong or outdated on this page?
26KUpdated 1 year agoApache-2.0
#Image-to-image#Inpainting#LoRA
FLUX.1 is a family of image models for people who want to generate or edit images on their own infrastructure. Black Forest Labs provides Python inference code for its open-weight models and a separate hosted API. Local inference runs on your hardware; API requests go to Black Forest Labs.
34.6KUpdated 1 day agoApache-2.0
macOS#ControlNet#Hugging Face integration#Image-to-image
4.6KUpdated 2 years agoApache-2.0
Web#ControlNet#Hugging Face integration#Image-to-image
13.2KUpdated 2 days agoApache-2.0
#ControlNet#Image-to-image#Inpainting
4.3KUpdated 10 months agoMIT
Web#Hugging Face integration#Image-to-image#LoRA
73.5KUpdated 4 years ago
#Guardrails#Hugging Face integration#Image-to-image
Stable Diffusion 1.5 is an AI image generation model for creators and developers who want to generate images on their own hardware. It turns text prompts into images and supports text-guided changes to existing pictures, including turning rough sketches into detailed artwork. Local inference keeps that image generation work on your machine.
Diffusers is an open-source Python library for developers and researchers who want to run diffusion models on their own hardware or build generation features into an application. It uses PyTorch and supports image, video and audio generation. The library is licensed under Apache 2.0 and supports Apple Silicon.
Kolors is a text-to-image model for people who want to generate photorealistic images on their own hardware, including work with Chinese prompts and Chinese cultural content. Developed by Kuaishou, it understands prompts in Chinese and English and can render text in both languages within generated images.
DiffSynth-Studio is a Python diffusion model engine for developers and researchers who want to generate media and train models on their own hardware. It supports large models on consumer GPUs through memory offloading and quantization, with inference and training in the same framework. It's open source under Apache 2.0.
OmniGen is a local AI image generation model that handles text prompts, reference images, and image editing within one model. It's for creators who want to reuse subjects across images and developers building image tools on their own hardware. The code is open source under the MIT license.