
OLMo is a family of language models for researchers and developers who want to inspect, adapt or run a model on their own hardware. Weights are downloadable. Ai2 also provides the training data, code, checkpoints and reports behind the models, giving researchers material to study the full training process rather than only the finished model.
The family includes Base models for further training, Think models that show intermediate reasoning steps, and Instruct models for chat, tool use and multi-turn dialogue. The smaller 7B models can run on a wider range of hardware; the 32B models include options aimed at demanding research and chat tasks. The site also offers a way to chat with OLMo, while downloadable artifacts support work on your own setup.
OLMo's research materials cover pretraining, later training stages and evaluation. The published data includes mixtures of web pages, code, books and scientific text, along with data used for instruction tuning where applicable. Researchers can use OlmoCore for training, Open Instruct for post-training and OLMES for reproducible evaluation. OLMoTrace connects model output to training data. The repository's code is licensed under Apache 2.0.
Claim this page with an email at allenai.org. OLMo gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find OLMo?Promote it
Something wrong or outdated on this page?
27.7KUpdated 9 months ago
#Batch processing#GGUF#Hugging Face integration
Qwen3 is a family of language models from Alibaba Cloud’s Qwen team for people who want to run models locally or on their own servers. It spans smaller and larger dense models as well as mixture-of-experts models. The weights are publicly available.
huggingface.coOpen-Weight LLMs
#Guardrails#Hugging Face integration#Multilingual
107Updated 1 year ago
#GGUF#Hugging Face integration#Multilingual
EXAONE 4.0 is a family of language models from LG AI Research that combines general language tasks and complex problem solving in the same model. It's aimed at developers building multilingual AI applications, including on-device apps and agents that use tools. It supports English, Korean and Spanish.
huggingface.coCoding Models
#Hugging Face integration#LoRA#Multilingual
20.4KUpdated 2 months agoApache-2.0
macOS · Linux#Hugging Face integration#LM Studio integration#Ollama integration
273Updated 2 years agoApache-2.0
Linux#Guardrails#Hugging Face integration#LM Studio integration
Command A is an open weights language model from Cohere and Cohere Labs for researchers and developers building self-hosted chatbots, document assistants and AI agents. Its focus is business tasks that combine multilingual text, supplied documents and external tools. You can run the model on your own hardware; Cohere also offers hosted chat through a playground and Hugging Face Space.
GLM-4.5 is an open-source language model for developers building AI agents and coding tools on their own servers. It combines reasoning with tool calling and offers a choice between thinking mode for complex tasks and non-thinking mode for direct responses. The MIT license permits commercial use and modification.
gpt-oss is a pair of OpenAI reasoning models for developers who want to run a local LLM or host one on their own server. The models are open weight and licensed under Apache 2.0. OpenAI also has a hosted browser demo, separate from running the models on your hardware.
Granite is IBM's family of open-source AI models for developers and businesses that want to run and customize AI on their own hardware or servers. The language-model repository listed here is archived and no longer maintained. The broader family includes models for language, speech, document understanding and forecasting, released under Apache 2.0 for research and commercial use.