macmon is a local system monitor for Apple Silicon Macs that reads hardware performance data without administrator privileges. It's for people checking resource use during local AI workloads or other demanding tasks, and developers who need those measurements in their own tools. It runs on macOS and supports M1 through M5 chips.
The terminal interface shows CPU, GPU and Apple Neural Engine power consumption alongside RAM and swap use, CPU and GPU temperatures, and fan speeds. Historical charts include averages and maximum values, so you can compare a brief spike with sustained load. The display can fit in a small terminal window.
CPU measurements distinguish time spent active from usage adjusted for clock frequency. That helps explain why a processor can be busy without using its full capacity. Detailed readings cover individual cores and the efficiency and performance clusters, with frequency and activity measurements for the GPU too.
For monitoring beyond the terminal, macmon exports JSON for scripts and serves JSON or Prometheus metrics over HTTP. It works with Prometheus and Grafana, and its metrics server can start automatically at login. Built-in stress tests generate CPU or GPU load, including cyclic CPU activity for checking how readings respond.
The project is open source under the MIT license and written in Rust. It reads the same performance data exposed by Apple's powermetrics through a private macOS API, without requiring root access. Developers can also use its Rust library to collect Apple Silicon metrics inside their own applications.
Claim this page and we'll verify you by hand. macmon gets the verified badge, and you can upgrade the listing to be featured on localhosted. Proud to be listed? Put our badge on your site.
Want more people to find macmon?Promote it
Something wrong or outdated on this page?
4.6KUpdated 4 years agoMIT
macOS
asitop is a terminal hardware monitor for Apple Silicon Macs, with separate views of CPU, GPU and Apple Neural Engine activity. Its README targets Apple Silicon on macOS Monterey; hardware-counter availability can differ on newer macOS releases. It runs locally and suits people who want to watch hardware use during demanding workloads, including local AI inference. It's free and open source under the MIT license.
2.3KUpdated 1 day agoMPL-2.0
macOS · Windows · Linux · Docker#Agent Skills#Batch processing#Multi-user access
223Updated 16 hours agoApache-2.0
macOS · Linux#GGUF#Git integration#Guardrails
11KUpdated 3 days ago
macOS · Windows · Linux · Docker
nvtop is a terminal monitor that shows activity across multiple GPUs and accelerators in an interface familiar to htop users. It's useful for people running local LLM workloads or managing compute servers who want to see which processes are using their hardware and how much GPU memory they consume. It runs locally on Linux and is open source under GPLv3 or later.
5.1KUpdated 20 hours agoApache-2.0
#Batch processing#Distributed execution#LoRA
6.9KUpdated 2 days agoApache-2.0
Docker · Web#Code execution#Git integration#Multi-user access
dstack is a self-hosted orchestration tool for AI teams managing compute across GPU clouds and their own servers. It puts cluster management, training jobs and model inference behind one interface, so teams can use different providers and accelerators without maintaining a separate workflow for each environment. It's open source under the Mozilla Public License 2.0.
LLMKube is a free, open-source Kubernetes operator for teams and homelab owners running local LLM inference across their own hardware. It manages Linux GPU servers and Apple Silicon Macs together, so a mixed fleet can serve models through the same platform. It uses the Apache 2.0 license.
AIBrix is open-source infrastructure for teams serving large language models on their own Kubernetes clusters. It focuses on the work around inference: directing requests, scaling capacity and managing models across servers. Enterprise infrastructure teams can use its components to build a self-hosted model service. It's licensed under Apache 2.0.
ClearML is an MLOps suite for recording experiments, managing datasets and running ML workloads. Its Apache 2.0 Python SDK connects to a ClearML Server, available as a hosted service or open-source software you deploy yourself. ClearML Agent handles job orchestration and reproducibility.