
Chroma is an 8.9B text-to-image model based on FLUX.1-schnell for creators and developers who want local image generation. The original model repository is deprecated; its maintainer recommends Chroma1-HD, Chroma1-Base or Chroma1-Flash instead. Those are separate successor checkpoints.
The original model card documents a ComfyUI workflow with a Chroma checkpoint, a T5 XXL text encoder and a FLUX VAE. Quantized alternatives can reduce memory requirements, but the exact setup depends on the selected checkpoint and runtime. The original card lists Diffusers support as work in progress, so use the documented ComfyUI route or check current runtime compatibility before deployment.
The weights are distributed in Safetensors format and use Apache 2.0, which permits use, modification and redistribution under its terms. The model is designed to generate unrestricted content and carries a sensitive-content designation; it may produce images unsuitable for all audiences.
Claim this page and we'll verify you by hand. Chroma gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Chroma?Promote it
Something wrong or outdated on this page?
1.1KUpdated 2 years agoApache-2.0
Web#Batch processing#Hugging Face integration#LoRA
CogView4 is a text-to-image model you can run on your own hardware, with support for Chinese and English prompts and Chinese text within generated images. It's aimed at developers and image creators who want local AI generation with native Chinese language support. The CogView4-6B model weights and repository code use Apache 2.0.
26KUpdated 1 year agoApache-2.0
#Image-to-image#Inpainting#LoRA
FLUX.1 is a family of image models for people who want to generate or edit images on their own infrastructure. Black Forest Labs provides Python inference code for its open-weight models and a separate hosted API. Local inference runs on your hardware; API requests go to Black Forest Labs.
2.5KUpdated 1 year agoMIT
Web#Hugging Face integration
HiDream-I1 is an open-source text-to-image model for people who want to generate images on their own hardware or build image generation into a Python application. It uses MIT licensing and supports local inference through CUDA, making it an option for developers and creators with NVIDIA GPU hardware.
17.8KUpdated 2 years agoMIT
Web#Batch processing#Hugging Face integration#Multimodal input
Janus-Pro is a multimodal AI model from DeepSeek that answers questions about images and creates pictures from text prompts. It runs on your own hardware and suits developers and researchers who want both capabilities in one model. A local Gradio demo provides a browser interface, while Hugging Face hosts a separate online demo.
4.6KUpdated 2 years agoApache-2.0
Web#ControlNet#Hugging Face integration#Image-to-image
Kolors is a text-to-image model for people who want to generate photorealistic images on their own hardware, including work with Chinese prompts and Chinese cultural content. Developed by Kuaishou, it understands prompts in Chinese and English and can render text in both languages within generated images.
1KUpdated 4 months agoApache-2.0
Linux · Web#Batch processing#Hugging Face integration#LoRA
Lumina-Image 2.0 is a local AI image generation framework for developers, researchers and people who want to generate images from text on their own hardware. It provides downloadable checkpoints, generation code and tools for adapting the model to your own image collections. The code uses the Apache 2.0 license.