
InternLM is a family of downloadable language models for developers and researchers building their own AI applications. It includes models for general conversation, complex reasoning and coding, with separate base and chat variants for customization or use in an assistant.
InternLM3-8B-Instruct supports ordinary conversational responses and a deep thinking mode that uses longer reasoning chains for difficult problems. InternLM2.5-Chat focuses on following instructions, conversation and function calling. Its Chat-1M variant accepts up to one million tokens of context, useful when an application needs to work with lengthy documents or large amounts of reference material. The family also supports internet search and combining information gathered across web pages; those tasks need internet access.
The models work with Hugging Face Transformers, with Python and PyTorch as part of the software stack. LMDeploy provides deployment and model serving within the surrounding toolchain. Developers can choose smaller models for lighter applications or larger variants for more complex tasks, rather than committing to one model size.
The repository uses the Apache 2.0 license. For teams adapting models, the ecosystem includes XTuner for training and fine-tuning, plus OpenCompass for evaluation. InternLM2-Reward provides separate models for assessing response preferences in reinforcement learning workflows.
Claim this page with an email at internlm.intern-ai.org.cn. InternLM gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find InternLM?Promote it
Something wrong or outdated on this page?
huggingface.coOpen-Weight LLMs
#Guardrails#Hugging Face integration#Multilingual
Command A is an open weights language model from Cohere and Cohere Labs for researchers and developers building self-hosted chatbots, document assistants and AI agents. Its focus is business tasks that combine multilingual text, supplied documents and external tools. You can run the model on your own hardware; Cohere also offers hosted chat through a playground and Hugging Face Space.
107Updated 1 year ago
#GGUF#Hugging Face integration#Multilingual
EXAONE 4.0 is a family of language models from LG AI Research that combines general language tasks and complex problem solving in the same model. It's aimed at developers building multilingual AI applications, including on-device apps and agents that use tools. It supports English, Korean and Spanish.
huggingface.coCoding Models
#Hugging Face integration#LoRA#Multilingual
20.4KUpdated 2 months agoApache-2.0
macOS · Linux#Hugging Face integration#LM Studio integration#Ollama integration
273Updated 2 years agoApache-2.0
Linux#Guardrails#Hugging Face integration#LM Studio integration
11.1KUpdated 11 months ago
#Hugging Face integration#Quantization#Tool calling
Kimi K2 is Moonshot AI's language model series for developers building coding assistants and AI agents, and researchers who want a foundation model to customize. You can run its checkpoints on your own infrastructure or use Moonshot's hosted API. Local inference runs on your hardware; the hosted API sends requests to Moonshot's service.
GLM-4.5 is an open-source language model for developers building AI agents and coding tools on their own servers. It combines reasoning with tool calling and offers a choice between thinking mode for complex tasks and non-thinking mode for direct responses. The MIT license permits commercial use and modification.
gpt-oss is a pair of OpenAI reasoning models for developers who want to run a local LLM or host one on their own server. The models are open weight and licensed under Apache 2.0. OpenAI also has a hosted browser demo, separate from running the models on your hardware.
Granite is IBM's family of open-source AI models for developers and businesses that want to run and customize AI on their own hardware or servers. The language-model repository listed here is archived and no longer maintained. The broader family includes models for language, speech, document understanding and forecasting, released under Apache 2.0 for research and commercial use.