Tools to Fine-Tune LLMs Locally

Fine-tune a language model with LoRA, QLoRA or full training on your GPUs, using Unsloth, LLaMA-Factory or TRL.

41 tools
Favicon of Reef

Reef

1 video
Self-hosted AI agent infrastructure that learns from feedback, trains weights with Slime and SGLang, or improves prompts and skills without local training GPUs.

7.4KUpdated 12 hours agoApache-2.0

Linux#Agent Skills#OpenAI-compatible API#Prompt versioning

An open-weight local LLM family from Google DeepMind for phones, PCs and servers, with Ollama and LM Studio support and an Apache 2.0 JAX library.

5.8KUpdated 2 days agoApache-2.0

Android#LM Studio integration#LoRA#Multilingual

Open-source LLM fine-tuning framework under Apache 2.0, with LoRA, QLoRA, multimodal training and inference through vLLM or SGLang.

75.2KUpdated 2 days agoApache-2.0

Web#LoRA#Multimodal input#OpenAI-compatible API

Favicon of Unsloth

Unsloth

6 videos
An open-source local LLM app for macOS, Windows and Linux. Run and train models, generate media, and connect coding agents to your hardware.

77KUpdated 21 hours agoApache-2.0

macOS · Windows · Linux · Docker · Web#Code execution#GGUF#Image-to-image

Open-source Python library for synthetic data generation and LLM training, with local models, API-based models, caching, and resumable workflows.

1.1KUpdated 2 years agoMIT

#Hugging Face integration#LoRA#Quantization

An open-source local LLM training toolkit in Python for training GPTs from scratch or fine-tuning GPT-2, with CPU, NVIDIA and Apple Silicon support.

63.5KUpdated 11 months agoMIT

macOS · Windows#Hugging Face integration

Open-source AI training framework built on PyTorch. Train on local CPUs or GPUs, fine-tune HuggingFace models, and serve models on your own server.

11.8KUpdated 4 days agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

An open-source LLM fine-tuning tool with a browser interface. Runs on Ubuntu with NVIDIA GPUs or in Docker, under the Apache 2.0 license.

5.2KUpdated 4 days agoApache-2.0

Linux · Docker · Web#Hugging Face integration#LoRA#Quantization

Apache 2.0 no-code model training tool for local hardware, Hugging Face Spaces or Colab. AutoTrain Advanced is no longer maintained.

4.6KUpdated 1 week agoApache-2.0

Web#Hugging Face integration#Multilingual

Self-hosted LLM gateway with evaluation, A/B testing, and Ollama support. Open source under Apache 2.0; archived and no longer maintained.

11.7KUpdated 4 months agoApache-2.0

Docker · Web#Batch processing#LLM tracing#Multimodal input

Self-hosted AI data assistant that queries databases, analyzes files, and generates reports with local models or cloud APIs. Open source under MIT.

20.1KUpdated 2 days agoMIT

macOS · Linux · Docker · Web#Code execution#llama.cpp backend#OpenAI-compatible API

A self-hosted Kubernetes operator that deploys Hugging Face models with vLLM, manages GPU capacity, and runs fine-tuning and document retrieval services.

1KUpdated 7 days ago

#Distributed execution#Hugging Face integration#LoRA

A family of AI models you can run offline with Ollama, llama.cpp or LM Studio, with open weights and training data for building specialized agents.

2.1KUpdated 3 weeks agoApache-2.0

Linux#GGUF#Guardrails#Hugging Face integration

A self-hosted vision-language model project for image chat, with a local Gradio interface, GPU inference and Apache 2.0 code.

25KUpdated 2 years agoApache-2.0

macOS · Web#LoRA#Multimodal input#Quantization

Code completion models that run locally on CPU or GPUs, with Hugging Face Transformers support and LoRA fine-tuning.

2.1KUpdated 3 years agoApache-2.0

#Hugging Face integration#LoRA#Quantization

Local LLM acceleration library for Intel CPUs, GPUs and NPUs. Runs on Windows and Linux, integrates with Ollama and llama.cpp, and is archived.

8.9KUpdated 8 months agoApache-2.0

Windows · Linux · Docker#Distributed execution#GGUF#Hugging Face integration

An open-source model quantization library for local LLMs and vision models, with Apache 2.0 licensing and Hugging Face Transformers integration.

960Updated 7 months agoApache-2.0

#Hugging Face integration#LoRA#Quantization

Open-source AI training and inference framework for your own GPU hardware, with distributed parallelism and Apache 2.0 licensing.

41.4KUpdated 3 days agoApache-2.0

#Distributed execution

Open-source LLM training engine for GPU and Ascend NPU hardware, with multimodal fine-tuning, reinforcement learning and Apache 2.0 licensing.

5.2KUpdated 6 days agoApache-2.0

Docker#Multimodal input

An open-source LLM training and deployment platform under Apache 2.0. Build specialized models on your own infrastructure or use its hosted service.

9.4KUpdated 2 days agoApache-2.0

Docker#Distributed execution#LoRA#Multimodal input

Open-source LLM training toolkit built on PyTorch. Train and chat with your own models on NVIDIA GPUs, with smaller CPU and Apple Silicon examples.

58.3KUpdated 3 months agoMIT

macOS#Code execution#Tool calling

An open-source LLM training toolkit that turns documents into specialist datasets, with offline generation on macOS and Linux and optional cloud compute.

1.9KUpdated 3 months agoMIT

macOS · Windows · Linux#Distributed execution#llama.cpp backend#Quantization

Self-hostable MLOps software for experiment tracking, pipelines, dataset versioning and model serving, with an Apache 2.0 Python SDK.

6.9KUpdated 2 days agoApache-2.0

Docker · Web#Code execution#Git integration#Multi-user access

Open-source distributed LLM software runs inference and fine-tuning across shared GPUs, with public or private networks and support for Llama 3.1.

10.6KUpdated 2 years agoMIT

macOS · Windows · Linux · Docker#Hugging Face integration

Open-source LLM training software for your own GPU servers, with PyTorch distributed training, NVIDIA and AMD support, and a BSD-3-Clause license.

5.8KUpdated 19 hours agoBSD-3-Clause

Linux#Distributed execution#Hugging Face integration

Local LLM desktop app for Windows, macOS and Linux. Run GGUF models offline or connect other apps through OpenAI- and Anthropic-compatible APIs.

47.7KUpdated 1 month agoAGPL-3.0

macOS · Windows · Linux · Docker · Web#GGUF#llama.cpp backend#LoRA

Favicon of TRL

TRL

1 video
Open-source Python library for LLM fine-tuning on your own GPUs, with Transformers support, preference training and Apache 2.0 licensing.

19.4KUpdated 1 day agoApache-2.0

#Distributed execution#LoRA#Quantization

An open-source local LLM training framework with LoRA, multimodal support, and deployment through vLLM, SGLang or LMDeploy. Apache 2.0 licensed.

15.8KUpdated 2 days agoApache-2.0

Web#Distributed execution#Hugging Face integration#LoRA

An open-source diffusion model training suite for Linux and Windows, with NVIDIA GPU support, LoRA training and a self-hosted web interface.

12.2KUpdated 3 days agoMIT

macOS · Windows · Linux · Web#Hugging Face integration#Image-to-image#LoRA

Open-source LLM training library for reinforcement learning and fine-tuning on NVIDIA, AMD and Ascend hardware, licensed under Apache 2.0.

23.7KUpdated 1 day agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input

Open-source Python toolkit for training and serving LLMs on your own hardware, with Apache 2.0 licensing, quantization and multi-GPU support.

13.7KUpdated 3 weeks agoApache-2.0

#Hugging Face integration#LoRA#Quantization

An open source RLHF framework for training models on your own NVIDIA GPUs, with HuggingFace model support and Ray, vLLM and DeepSpeed backends.

10.1KUpdated 2 weeks agoApache-2.0

Docker#Distributed execution#Hugging Face integration#LoRA

An open-source LLM fine-tuning library for single-GPU and distributed training, with LoRA and QLoRA support. Development is no longer active.

5.8KUpdated 5 months agoBSD-3-Clause

#Hugging Face integration#LoRA#Quantization

Favicon of DSPy

DSPy

2 videos
Open-source Python framework for building modular LLM systems, with typed outputs, tool-using agents, and automatic prompt optimization. MIT licensed.

38.4KUpdated 4 days agoMIT

#Code execution#MCP#Multimodal input

A PyTorch quantization library that reduces LLM memory use for inference and fine-tuning with 8-bit optimizers, LLM.int8() and QLoRA. MIT licensed.

8.5KUpdated 4 weeks agoMIT

macOS · Windows · Linux#LoRA#Quantization

An open-source local LLM framework that splits work across CPUs and GPUs, with SGLang serving and LlamaFactory fine-tuning under Apache 2.0.

19.5KUpdated 1 week agoApache-2.0

Docker#LoRA#Multimodal input#Prompt caching

More in Fine-Tuning and Training