Open-Source Training and RLHF Frameworks

Frameworks for pretraining, distributed training and reinforcement learning, including DeepSpeed, verl and nanoGPT.

45 tools
Favicon of Reef

Reef

1 video
Self-hosted AI agent infrastructure that learns from feedback, trains weights with Slime and SGLang, or improves prompts and skills without local training GPUs.

7.4KUpdated 12 hours agoApache-2.0

Linux#Agent Skills#OpenAI-compatible API#Prompt versioning

Favicon of MLX

MLX

2 videos
An open-source machine learning array framework with NumPy-style APIs, shared CPU and GPU memory on Apple silicon, and Linux CPU and CUDA backends.

28.6KUpdated 24 hours agoMIT

macOS · Linux#Distributed execution#LoRA

Open-source LLM fine-tuning framework under Apache 2.0, with LoRA, QLoRA, multimodal training and inference through vLLM or SGLang.

75.2KUpdated 2 days agoApache-2.0

Web#LoRA#Multimodal input#OpenAI-compatible API

An open-source Python library for running and training text, vision, audio and multimodal models locally, with Apache 2.0 licensing and PyTorch support.

166.8KUpdated 1 day agoApache-2.0

#Hugging Face integration#Multimodal input

Open-source Python library for synthetic data generation and LLM training, with local models, API-based models, caching, and resumable workflows.

1.1KUpdated 2 years agoMIT

#Hugging Face integration#LoRA#Quantization

An open-source singing voice conversion framework that runs fully offline with user-trained models. Licensed under AGPL-3.0; archived and no longer maintained.

28.1KUpdated 3 years agoAGPL-3.0

#Hugging Face integration#ONNX#Voice conversion

An open-source local LLM training toolkit in Python for training GPTs from scratch or fine-tuning GPT-2, with CPU, NVIDIA and Apple Silicon support.

63.5KUpdated 11 months agoMIT

macOS · Windows#Hugging Face integration

Open-source AI training framework built on PyTorch. Train on local CPUs or GPUs, fine-tune HuggingFace models, and serve models on your own server.

11.8KUpdated 4 days agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

Apache 2.0 no-code model training tool for local hardware, Hugging Face Spaces or Colab. AutoTrain Advanced is no longer maintained.

4.6KUpdated 1 week agoApache-2.0

Web#Hugging Face integration#Multilingual

A Python library for running and training CLIP image-text models on your own hardware, with local checkpoints and Hugging Face model support.

14.2KUpdated 5 days ago

#Hugging Face integration#Multimodal input

Open-source PyTorch training framework for pretraining and fine-tuning models on your own CPUs or GPUs, with Apache 2.0 licensing.

31.4KUpdated 1 week agoApache-2.0

macOS#Distributed execution#ONNX

Open-source LLM pretraining library built on PyTorch for custom datasets, with distributed NVIDIA GPU training. Licensed under Apache 2.0.

2.8KUpdated 1 week agoApache-2.0

Linux#Distributed execution#Hugging Face integration

Open-source AI training and inference framework for your own GPU hardware, with distributed parallelism and Apache 2.0 licensing.

41.4KUpdated 3 days agoApache-2.0

#Distributed execution

An open-source diffusion model trainer that splits large models across GPUs. Uses DeepSpeed and supports Windows through WSL 2 under GPL-3.0.

2KUpdated 2 days agoGPL-3.0

Windows · Linux#Distributed execution#LoRA#Quantization

A text-to-music model with melody conditioning and local GPU inference. AudioCraft code is MIT licensed; pretrained weights have a noncommercial license.

23.7KUpdated 2 years agoMIT

#Multimodal input

A self-hosted diffusion model trainer with a web UI, LoRA and full fine-tuning, NVIDIA, AMD and Apple Silicon support, and an AGPL-3.0 license.

2.9KUpdated 19 hours agoAGPL-3.0

macOS · Docker · Web#ControlNet#Distributed execution#Human approval

Open-source LLM training engine for GPU and Ascend NPU hardware, with multimodal fine-tuning, reinforcement learning and Apache 2.0 licensing.

5.2KUpdated 6 days agoApache-2.0

Docker#Multimodal input

An open-source LLM training and deployment platform under Apache 2.0. Build specialized models on your own infrastructure or use its hosted service.

9.4KUpdated 2 days agoApache-2.0

Docker#Distributed execution#LoRA#Multimodal input

Open-source Python tools optimize Hugging Face models for local, edge and cloud hardware, with ONNX Runtime, OpenVINO and TensorRT-LLM integrations.

3.5KUpdated 6 days agoApache-2.0

#Hugging Face integration#ONNX#Quantization

Open-source Python library for object detection and segmentation, with pretrained models and deployment exports. Licensed under Apache 2.0.

34.7KUpdated 2 days agoApache-2.0

AI model hub with an Apache 2.0 Python library for local inference, training and evaluation, plus hosted demos and cloud notebooks.

9.2KUpdated 6 days agoApache-2.0

Docker#Image-to-image#Inpainting#Multimodal input

Open-source LLM training toolkit built on PyTorch. Train and chat with your own models on NVIDIA GPUs, with smaller CPU and Apple Silicon examples.

58.3KUpdated 3 months agoMIT

macOS#Code execution#Tool calling

An open-source PyTorch toolkit for training and testing object detection and segmentation models, with GPU operations and an Apache 2.0 license.

33KUpdated 3 years agoApache-2.0

Open-source LLM training software for your own GPU servers, with PyTorch distributed training, NVIDIA and AMD support, and a BSD-3-Clause license.

5.8KUpdated 19 hours agoBSD-3-Clause

Linux#Distributed execution#Hugging Face integration

PyTorch quantization library that reduces model memory use and speeds training and inference on your hardware, with CPU, GPU and mobile deployment support.

3KUpdated 5 days ago

Linux · iOS#Hugging Face integration#LoRA#Quantization

Favicon of DeepSpeed

DeepSpeed

1 video
Open-source PyTorch optimization library for distributed training and inference, with Apache 2.0 licensing and support for NVIDIA and AMD GPUs.

43.2KUpdated 1 day agoApache-2.0

Windows#Distributed execution

An open source engine for running ONNX models on Windows, macOS, Linux, mobile devices and the web, with CPU, GPU and NPU support.

21.9KUpdated 1 day agoMIT

macOS · Windows · Linux · iOS · Android · Web#Distributed execution#ONNX

Favicon of TRL

TRL

1 video
Open-source Python library for LLM fine-tuning on your own GPUs, with Transformers support, preference training and Apache 2.0 licensing.

19.4KUpdated 1 day agoApache-2.0

#Distributed execution#LoRA#Quantization

An open-source local LLM training framework with LoRA, multimodal support, and deployment through vLLM, SGLang or LMDeploy. Apache 2.0 licensed.

15.8KUpdated 2 days agoApache-2.0

Web#Distributed execution#Hugging Face integration#LoRA

An open-source diffusion model training suite for Linux and Windows, with NVIDIA GPU support, LoRA training and a self-hosted web interface.

12.2KUpdated 3 days agoMIT

macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#LoRA

A vision-language model family with local Python code for training and evaluation, Apache 2.0 licensing, and variants built on OLMo and Qwen2.

937Updated 2 years agoApache-2.0

#Hugging Face integration#Multimodal input#Works offline

Open-source LLM training library for reinforcement learning and fine-tuning on NVIDIA, AMD and Ascend hardware, licensed under Apache 2.0.

23.7KUpdated 1 day agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input

Open-source Python toolkit for training and serving LLMs on your own hardware, with Apache 2.0 licensing, quantization and multi-GPU support.

13.7KUpdated 3 weeks agoApache-2.0

#Hugging Face integration#LoRA#Quantization

An open source RLHF framework for training models on your own NVIDIA GPUs, with HuggingFace model support and Ray, vLLM and DeepSpeed backends.

10.1KUpdated 2 weeks agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

An open-source LLM fine-tuning library for single-GPU and distributed training, with LoRA and QLoRA support. Development is no longer active.

5.8KUpdated 5 months agoBSD-3-Clause

#Hugging Face integration#LoRA#Quantization

Open-source Python toolkit for local AI audio generation, fine-tuning and training, with a Gradio interface and support for Stable Audio Open.

3.9KUpdated 4 months agoMIT

Web#Hugging Face integration

More in Fine-Tuning and Training