TagGUI is a desktop app for people preparing image datasets for generative AI training. It combines local AI captioning with manual tag editing, so you can generate descriptions and correct them in the same workspace. It's open source under GPL-3.0 and runs on Windows, Linux and macOS, though macOS doesn't have a packaged release.
Caption generation runs on your hardware, using a CPU or a compatible NVIDIA GPU. The app can download captioning models and use models you've already stored locally. It generates captions or tags for multiple images at once, with controls for prompts, caption prefixes and preferred or discouraged wording. Prompts can draw on an image's existing tags, filename or folder name; wording requests aren't guaranteed to appear in the output.
The editing tools focus on repeated work across a dataset. Keyboard shortcuts and autocomplete based on your most-used tags reduce typing, while batch operations let you rename, delete, sort or reorder tags across images. A built-in Stable Diffusion token counter helps you assess caption length.
Filters let you find images by tags, caption text, filenames or paths, as well as by tag count, character count and token count. You can combine conditions or use wildcards to isolate groups that need attention. TagGUI reads tags from matching text files beside your images and automatically saves edits back to those files.
Claim this page and we'll verify you by hand. TagGUI gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find TagGUI?Promote it
Something wrong or outdated on this page?
89Updated 22 hours agoGPL-3.0
macOS · Windows · Linux · Docker · Web#Batch processing#Hugging Face integration#ONNX
PixlStash is an open-source image manager for photographers, AI creators and people curating image datasets. It combines library search with tools for reviewing tags, ranking images and sending selected work through ComfyUI. You can use a desktop app or a self-hosted server with a browser interface.
1.3KUpdated 7 months agoApache-2.0
Windows · Docker#Hugging Face integration#Multimodal input#OpenAI-compatible API
1KUpdated 7 days agoMIT
macOS · Windows · Linux#Batch processing#Multimodal input#Semantic search
808Updated 7 years agoGPL-2.0
macOS · Windows · Linux#Batch processing
922Updated 5 months agoMIT
macOS · Windows · Linux · Docker · Web#llama.cpp backend#LM Studio integration#MLX
16.2KUpdated 18 hours agoGPL-3.0
macOS · Windows · Linux#Multimodal input
LabelMe is a desktop image annotation app for people preparing computer vision datasets. It combines manual drawing with AI assistance for outlining objects and creating labels from text. It runs on 64-bit macOS, Windows and Linux..
JoyCaption is an open-weight image captioning model for people preparing datasets to train or fine-tune diffusion models. It runs on your own GPU and covers both SFW and NSFW images, including photography, anime, digital art and furry artwork. Automated captions reduce the need to write descriptions by hand or find images that already have usable text.
rclip searches image folders by their visual content, so you can find photos without adding tags or importing them into a photo library. It's a local AI tool for people who keep image collections on their own computers or servers and prefer working in the terminal. It runs on Linux, Windows and Apple Silicon macOS, and it's open source under the MIT license.
digiKam is a local photo manager and image editor for photographers who need to organize large collections of photos, RAW files and videos. It runs on Linux, Windows and macOS and is open-source software under GPL-2.0. Its archived GitHub mirror points contributors to the maintained KDE repository.
Eclaire is a self-hosted AI assistant for notes, documents, photos, bookmarks and tasks. You can ask questions about saved material, inspect the sources behind an answer and ask the assistant to create notes or update tasks. Scheduled automations handle recurring requests such as a weekly task summary.