
Ludwig is an open-source Python framework for developers and researchers who want to train custom AI models on their own hardware. A YAML file describes the model and training pipeline, while Ludwig handles preprocessing, training and evaluation. It uses the Apache 2.0 license. Install the Python package with the optional LLM dependencies for fine-tuning; current source requires Python 3.12 or later.
Built on PyTorch, it supports local CPU or GPU training and distributed jobs through Ray, DeepSpeed and FSDP. Docker images cover CPU, GPU and Ray deployments, and KubeRay supports Kubernetes clusters. The same model definition can carry across local and distributed training.
For LLM fine-tuning, Ludwig works with HuggingFace models including Llama, Mistral and Qwen. It supports instruction tuning and preference alignment methods such as DPO and GRPO. LoRA and 4-bit QLoRA reduce training memory needs, with single-GPU QLoRA workflows when the model fits the available memory. Vision-language models such as LLaVA and Qwen2-VL are also supported.
Its scope extends beyond LLMs. Models can combine text with images, audio, tabular data or time series and learn multiple outputs together. Applications include classification, image segmentation, speech recognition and forecasting. Custom encoders, losses and metrics give developers room to extend the framework.
AutoML can find an initial baseline, while Optuna and Ray Tune search training settings. Feature importance and explainability tools help inspect results. Experiment tracking connects to MLflow, TensorBoard and Weights & Biases. Trained models can run behind a self-hosted REST API or export to SafeTensors, ONNX and torch.export.
Claim this page with an email at ludwig.ai. Ludwig gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find Ludwig?Promote it
Something wrong or outdated on this page?
12.5KUpdated 2 days agoApache-2.0
Docker#Distributed execution#Hugging Face integration#LoRA
Axolotl is an open-source LLM fine-tuning framework for developers, researchers, and teams training models on their own data. It runs on local hardware or cloud infrastructure you control, including Docker and Kubernetes environments. The framework uses Apache 2.0, which permits commercial use.
10.1KUpdated 2 weeks agoApache-2.0
Docker#Distributed execution#Hugging Face integration#LoRA
9.4KUpdated 2 days agoApache-2.0
Docker#Distributed execution#LoRA#Multimodal input
5.2KUpdated 6 days agoApache-2.0
Docker#Multimodal input
XTuner is an open-source LLM training engine for researchers and teams training large mixture-of-experts (MoE) models on their own hardware. It supports GPU and Ascend NPU training, with an emphasis on memory use and distributed training efficiency at scales reaching a trillion parameters.
12.2KUpdated 3 days agoMIT
macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#LoRA
1.1KUpdated 2 years agoMIT
#Hugging Face integration#LoRA#Quantization
OpenRLHF is a self-hosted Python framework for researchers and teams training language models with human feedback or custom rewards. It runs on your own NVIDIA GPU hardware, with Docker support and distributed training across servers. It's open source under Apache 2.0.
Oumi builds specialized AI models for teams that want control over their training data, model weights, and deployment. Its Apache 2.0 open-source stack runs on laptops, clusters, and your own servers, while its hosted service automates model development from a plain-English task description. You own the resulting weights, data, and training recipes.
AI Toolkit (ostris) is an MIT-licensed training suite for people who want to fine-tune image and video models on their own hardware or a self-hosted server. It targets consumer NVIDIA GPUs and runs on Linux and Windows, including ARM64 Linux systems such as DGX Spark. An experimental installer also supports Apple Silicon Macs. GPU memory needs depend on the model and training task.
DataDreamer connects LLM prompting, synthetic data generation, and model training in one Python library. It's for researchers and developers who want to build datasets and use them to fine-tune or align models in reproducible workflows. The library is open source under the MIT license.