Lumina-Image 2.0 is a local AI image generation framework for developers, researchers and people who want to generate images from text on their own hardware. It provides downloadable checkpoints, generation code and tools for adapting the model to your own image collections. The code uses the Apache 2.0 license.
ComfyUI support makes it an option for visual image workflows, while the Diffusers integration suits Python applications. A Gradio interface provides browser-based access to a locally running model. Batch generation handles collections of prompts, so you can produce images without submitting each request individually. An official Hugging Face Space also offers a hosted demo; that demo runs remotely rather than on your machine.
Fine-tuning support includes training on image and text pairs and LoRA adaptation through Diffusers. The companion Lumina-Accessory project builds on the model for controllable generation, image editing and identity preservation, with support for training on one task or several tasks together.
The model uses Gemma-2-2B to encode text and FLUX-VAE-16CH for image encoding and decoding. Access to Gemma requires Hugging Face authentication. Local generation can use downloaded checkpoints, including .pth weights. The Diffusers integration supports CPU offloading to reduce GPU memory use.
Claim this page and we'll verify you by hand. Lumina-Image 2.0 gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Lumina-Image 2.0?Promote it
Something wrong or outdated on this page?
165.2KUpdated 2 years agoAGPL-3.0
macOS · Windows · Linux · Web#Batch processing#Code execution#Image-to-image
Stable Diffusion web UI (AUTOMATIC1111) is a browser interface for generating and editing images with models running on your own hardware. It's for artists and anyone who wants control over prompts, models and image variations. The software is open source under AGPL-3.0.
4.6KUpdated 2 years agoApache-2.0
Web#ControlNet#Hugging Face integration#Image-to-image
4.3KUpdated 10 months agoMIT
Web#Hugging Face integration#Image-to-image#LoRA
575Updated 24 hours agoGPL-3.0
macOS · Linux · iOS · Docker#Image-to-image#Inpainting#LoRA
2.1KUpdated 3 days ago
Windows · Linux#LoRA#Quantization
Musubi Tuner is a Python toolkit for training LoRA adapters for image and video generation models on your own hardware. It's aimed at people who want to customize these models using their own datasets and are comfortable working with training scripts. It also includes image and video generation scripts for supported architectures.
1.9KUpdated 2 years agoApache-2.0
Web#Hugging Face integration
PixArt-Sigma is a text-to-image diffusion model that supports generation at 2K and 4K resolutions on your own machine or server. It's aimed at developers and researchers who want pretrained models they can run themselves, along with code for training and adapting them. The Python code is open source under Apache 2.0.
Kolors is a text-to-image model for people who want to generate photorealistic images on their own hardware, including work with Chinese prompts and Chinese cultural content. Developed by Kuaishou, it understands prompts in Chinese and English and can render text in both languages within generated images.
OmniGen is a local AI image generation model that handles text prompts, reference images, and image editing within one model. It's for creators who want to reuse subjects across images and developers building image tools on their own hardware. The code is open source under the MIT license.
Draw Things is an AI image generation app for iPhone, iPad and Mac that keeps generation on your device and works offline. It's for people who want to create and edit images without sending that work to a cloud service, including artists developing character concepts or trying out apparel designs.