mergekit combines existing language models into a single model on your own hardware, without additional training or access to the original training data. It's for developers and researchers who want to combine fine-tuned capabilities or adjust the balance between model behaviors. The Python toolkit is open source under LGPL-3.0.
Merges can run entirely on CPU, or use GPU acceleration with as little as 8 GB of VRAM. It loads model weights only as needed, reducing memory demands when the input models don't fit in memory at once. Supported model families include Llama, Mistral, GPT-NeoX and StableLM.
Its methods include weighted averaging, SLERP, TIES, DARE and Arcee Fusion, giving users different ways to combine checkpoints and handle conflicting changes. It can also assemble models from selected layers, build mixture-of-experts models from dense models, and chain merges so one result becomes the input to another.
Beyond merging, mergekit can extract PEFT-compatible LoRA adapters from fine-tuned models. Tokenizer controls align vocabularies across inputs, while tokenizer transplantation supports draft models for speculative decoding. For work outside Hugging Face Transformers, it applies merge algorithms to raw PyTorch .pt and .safetensors checkpoints.
The toolkit runs locally. FrankensteinAI is a separate hosted service powered by mergekit, with a browser interface and a community gallery for sharing and comparing merged models.
Claim this page and we'll verify you by hand. mergekit gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find mergekit?Promote it
Something wrong or outdated on this page?
1.1KUpdated 2 years agoMIT
#Hugging Face integration#LoRA#Quantization
DataDreamer connects LLM prompting, synthetic data generation, and model training in one Python library. It's for researchers and developers who want to build datasets and use them to fine-tune or align models in reproducible workflows. The library is open source under the MIT license.
6.3KUpdated 10 months agoApache-2.0
#Hugging Face integration#llama.cpp backend#LoRA
2.4KUpdated 2 years agoAGPL-3.0
macOS · Windows · Linux · Docker · Web#Hugging Face integration
62.2KUpdated 1 month agoMIT
macOS · Windows · Linux · Docker · Web#Hugging Face integration#Multilingual#Voice activity detection
8.7KUpdated 2 years agoMIT
Linux#Hugging Face integration#Multimodal input#ONNX
26.6KUpdated 1 day agoApache-2.0
Docker#Guardrails#Hugging Face integration#Hybrid search
Orpheus TTS is an open-source text-to-speech system for developers building voice applications or adapting speech models to their own recordings. It runs locally and uses a Llama backbone to generate speech with control over emotion and intonation. The code uses the Apache 2.0 license.
AllTalk TTS generates speech on your own computer. The project recommends v2 for most users; the saved documentation below describes v1, built on Coqui TTS and XTTSv2 models. It's for people adding voices to AI conversations or producing spoken audio from longer texts. It runs as a standalone application or alongside Text-generation-webui, with support for Windows, Linux and macOS.
GPT-SoVITS is a local text-to-speech and voice cloning tool. It can generate speech from a short reference recording or fine-tune a model for a custom voice. The source code uses the MIT license.
Hallo turns a single portrait and a speech recording into an animated talking video on your own hardware. It's a local AI tool for creators working with talking portraits and researchers who want access to both generation and training code. The Python code uses the MIT license; required pretrained models and dependencies have their own terms.
Haystack is a Python framework for developers building self-hosted AI agents, document search, and apps that answer questions using their own data. Its modular pipelines let teams control which information reaches a model and inspect how retrieval, memory, tools, and generation contribute to an answer. It's open source under Apache 2.0.