Open Text-to-Image Models

Open-weight models that generate or edit images from a text prompt, such as FLUX.1, Stable Diffusion XL and Qwen-Image.

14 tools
A local text-to-image model for Chinese and English prompts, with Diffusers support, a community ComfyUI wrapper, and Apache 2.0 code.

1.1KUpdated 2 years agoApache-2.0

Web#Batch processing#Hugging Face integration#LoRA

An Apache 2.0 text-to-image model for local ComfyUI workflows. Its original repository is deprecated in favor of newer Chroma1 checkpoints.

huggingface.coImage Generation Models

An image generation model you can run locally, with text-to-image and image-to-image support, MIT-licensed Python code and separately licensed weights.

27.3KUpdated 9 months agoMIT

Web#Hugging Face integration#Image-to-image

Favicon of Qwen-Image

Qwen-Image

1 video
An open-source image generation and editing model for local deployment, with Chinese text rendering, multi-image edits and an Apache 2.0 license.

8.4KUpdated 8 months agoApache-2.0

Web#Image-to-image#LoRA#Multimodal input

A local AI image generation model with text-guided image editing, Diffusers support, and a reference implementation requiring at least 10GB of GPU VRAM.

73.5KUpdated 4 years ago

#Guardrails#Hugging Face integration#Image-to-image

A local image model with Chinese and English prompts, text rendering, ComfyUI support, Apache 2.0 code, and separate model-weight terms.

4.6KUpdated 2 years agoApache-2.0

Web#ControlNet#Hugging Face integration#Image-to-image

An open-source text-to-image model for local or self-hosted generation, with 4K output, PyTorch training code and Hugging Face Diffusers support.

1.9KUpdated 2 years agoApache-2.0

Web#Hugging Face integration

An open-source image and video generation framework with 4K text-to-image models, laptop GPU support, and ComfyUI and Diffusers integrations.

9.2KUpdated 2 weeks agoApache-2.0

#ControlNet#LoRA#Multimodal input

An MIT-licensed image generation model you can run locally with CUDA, with full and distilled variants, Diffusers support, and a Gradio interface.

2.5KUpdated 1 year agoMIT

Web#Hugging Face integration

An open-source image generation model that runs locally, edits images, and uses multiple references. Supports Diffusers and carries an MIT license.

4.3KUpdated 10 months agoMIT

Web#Hugging Face integration#Image-to-image#LoRA

A local multimodal AI model for answering image questions and generating pictures, with downloadable weights and a Gradio interface.

17.8KUpdated 2 years agoMIT

Web#Batch processing#Hugging Face integration#Multimodal input

An open-source image generation framework for local use, with ComfyUI, Diffusers and fine-tuning support. Code uses the Apache 2.0 license.

1KUpdated 4 months agoApache-2.0

Linux · Web#Batch processing#Hugging Face integration#LoRA

A text-to-image model family you can deploy on your own infrastructure, with a consumer-hardware variant and separate API and web access.

1.5KUpdated 2 years agoMIT

#ControlNet#Hugging Face integration#Image-to-image

Local image generation and editing with FLUX.1 models. Run open weights on your own infrastructure or use Black Forest Labs' hosted API.

26KUpdated 1 year agoApache-2.0

#Image-to-image#Inpainting#LoRA

More in Open Models