gpumod is an open-source command-line tool for managing GPU services tailored to machine learning workloads. It supports integration with tools like llama-cpp and vllm, helping ML engineers efficiently allocate and monitor GPU resources.
gpumod is an Inference & model serving project. It focuses on managing GPU resources and services for machine learning workloads efficiently. gpumod is an open-source project aimed at ML engineers and researchers. The project is open source (Apache-2.0). It ships for the command line, and it can be self-hosted.
jaigouk builds and maintains gpumod, and it first shipped in 2026. Development happens publicly on GitHub with 13 stars and 111 commits in the last 90 days. Key capabilities include GPU service management, ML workload support, and llama-cpp integration.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do