GPU Monitoring and Management Tools

Watch GPU load, memory and temperature or share GPUs between containers, with nvtop, nvitop, asitop and the NVIDIA Container Toolkit.

17 tools
Self-hosted LLM inference for Kubernetes with NVIDIA, AMD and Apple Silicon support, OpenAI-compatible APIs, and an Apache 2.0 license.

223Updated 16 hours agoApache-2.0

macOS · Linux#GGUF#Git integration#Guardrails

An open-source GPU computing platform for AI training and inference on AMD hardware, with Linux and Windows support and PyTorch, JAX, vLLM and SGLang.

6.8KUpdated 4 weeks agoMIT

Windows · Linux

A terminal performance monitor for Apple Silicon Macs that tracks CPU, GPU, memory and power use. Runs locally on macOS and uses the MIT license.

4.6KUpdated 4 years agoMIT

macOS

A self-hosted Kubernetes operator that deploys Hugging Face models with vLLM, manages GPU capacity, and runs fine-tuning and document retrieval services.

1KUpdated 7 days ago

#Distributed execution#Hugging Face integration#LoRA

Open-source AI compute management software that runs jobs across your Kubernetes and Slurm clusters or cloud accounts, using GPUs, TPUs and CPUs.

10.7KUpdated 21 hours agoApache-2.0

#Code execution#Distributed execution#Multi-user access

Self-hosted LLM observability and evaluation platform under Apache 2.0. Monitor Ollama, vLLM and coding agents with portable OpenTelemetry traces.

2.8KUpdated 1 day agoApache-2.0

Windows · Linux · Docker · Web#LLM tracing#Ollama integration#Prompt versioning

Open-source LLM serving infrastructure for Kubernetes with multi-node inference, demand-based autoscaling, LoRA management and vLLM integration.

5.1KUpdated 20 hours agoApache-2.0

#Batch processing#Distributed execution#LoRA

Self-hostable MLOps software for experiment tracking, pipelines, dataset versioning and model serving, with an Apache 2.0 Python SDK.

6.9KUpdated 2 days agoApache-2.0

Docker · Web#Code execution#Git integration#Multi-user access

Self-hosted AI compute orchestration under MPL-2.0 for training and inference on GPU clouds, Kubernetes, VMs and bare-metal servers.

2.3KUpdated 1 day agoMPL-2.0

macOS · Windows · Linux · Docker#Agent Skills#Batch processing#Multi-user access

A command-line GPU monitor that runs locally on NVIDIA hardware, shows memory use and running processes, and provides JSON output under the MIT license.

4.4KUpdated 2 weeks agoMIT

A local system monitor for Apple Silicon Macs that tracks power, temperature and memory without root access, with JSON and Prometheus output. It's MIT licensed.

1.9KUpdated 2 months agoMIT

macOS

An open-source container toolkit that gives Docker workloads access to NVIDIA GPUs on Linux. Requires the NVIDIA driver, but not the host CUDA Toolkit.

4.6KUpdated 1 week agoApache-2.0

Linux

GPU management and monitoring software for NVIDIA data-center GPUs on Linux, with Kubernetes telemetry and an Apache 2.0 open-source core.

798Updated 1 month agoApache-2.0

Linux

A local NVIDIA Jetson monitoring tool with a terminal interface, Python API and Docker support. Open source under AGPL-3.0.

2.6KUpdated 1 week agoAGPL-3.0

Linux · Docker

An open-source NVIDIA GPU process monitor for Linux and Windows, with interactive terminal views, Python APIs and Grafana dashboard support.

7.2KUpdated 1 day agoApache-2.0

Windows · Linux · Docker · Web

A self-hosted inference framework that coordinates NVIDIA GPU clusters with vLLM, SGLang or TensorRT-LLM and exposes an OpenAI-compatible API.

8.2KUpdated 20 hours ago

#Distributed execution#Multimodal input#OpenAI-compatible API

A local GPU and accelerator monitor for Linux, with process lists, usage charts and support for NVIDIA, AMD, Intel and dedicated AI hardware.

11KUpdated 3 days ago

macOS · Windows · Linux · Docker

More in GPUs, Clusters and Edge