mods is a command-line AI tool for people who want to ask questions about command output or use model responses in shell pipelines. It connects to local LLMs through LocalAI as well as cloud services. The project is archived and no longer maintained.
Its main use is working with text already in the terminal. It can combine command output with a prompt, accept a prompt on its own, and request responses in Markdown, JSON or other text formats. That makes it useful for developers and shell users who want AI responses alongside their existing commands rather than in a separate chat window.
The client runs on macOS, Linux and Windows, with FreeBSD support through ports. Model processing happens at the endpoint you choose: LocalAI runs models locally, while OpenAI, Azure OpenAI, Cohere, Groq and Google Gemini receive input through their cloud APIs. Those cloud connections require API keys. mods also works with OpenAI-compatible endpoints, and its documented LocalAI support includes GPT4ALL-J.
Conversations stay local by default. You can return to a saved conversation, continue a previous exchange or turn off conversation saving. Custom roles let you reuse system prompts for recurring tasks, such as requesting shell commands with minimal explanation. MCP support lets it connect to servers that expose tools to the model.
mods is open source under the MIT license. Charm's Crush includes much of its functionality in a non-interactive mode.
Claim this page and we'll verify you by hand. mods gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find mods?Promote it
Something wrong or outdated on this page?
10.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux · Android · Web#Code execution#MCP#Multimodal input
aichat brings Ollama and cloud AI services into the same terminal interface for developers and people who work at the command line. It runs locally on macOS, Linux and Windows, with Android support through Termux. Model processing happens through the backend you choose: Ollama supports local models, while providers such as OpenAI, Claude and Gemini process requests in the cloud.
12.3KUpdated 3 months agoMIT
macOS · Windows · Linux · Docker#Code execution#Git integration#Human approval
250.1KUpdated 18 hours agoMIT
macOS · Windows · Linux · Android · Docker#Agent Skills#Code execution#Multi-agent workflows
2.2KUpdated 3 days agoMIT
macOS · Windows · Linux#Batch processing#GGUF#Guardrails
1.5KUpdated 7 months agoApache-2.0
macOS · Windows · Linux#Ollama integration#RAG#Works offline
3.1KUpdated 3 months agoMIT
macOS · Windows · Linux#Distributed execution#Hugging Face integration#Quantization
ShellGPT is an AI terminal assistant for developers and people who work with shell commands. It runs on Linux, macOS and Windows, turning plain-language requests into commands suited to your operating system and shell. You can review, explain or execute its suggestions, and its Bash and Zsh integrations put generated commands into the terminal input line for editing.
Hermes Agent is a self-hosted AI agent from Nous Research for people who want an assistant that remembers past work and develops reusable skills. It creates skills after complex tasks, revises them through use, and searches earlier conversations to recover context across sessions. The project is open source under the MIT license.
node-llama-cpp is an open source library for developers adding local LLM inference to JavaScript and TypeScript applications. It connects Node.js, Bun and Electron to llama.cpp, running GGUF models on your own machine. Its MIT license allows use in commercial projects.
tlm is an open source terminal assistant for people who want help writing shell commands and understanding unfamiliar ones. It runs models on your workstation through Ollama, so command assistance doesn't depend on a cloud service. It works offline and requires no API key or subscription.
Distributed Llama runs a local LLM across several computers, sharing both the computation and the model's memory use. It's for people who want to use their own networked hardware for inference rather than keep the entire workload on one machine. The C++ project is open source under the MIT license.