inferbench-cli is an open-source CLI tool for benchmarking local LLM inference speed and providing hardware configuration advice for omlx and llama.cpp. It helps developers and researchers optimize AI model performance on their own machines.
inferbench-cli sits in PulseGate's LLM evaluation & benchmarks category. It focuses on measuring and optimizing local LLM inference performance on user hardware. inferbench-cli is an open-source project aimed at AI developers and researchers. The project is open source (Apache-2.0). It ships for the command line, and it can be self-hosted.
Behind inferbench-cli is Rudrendu Paul, and it first shipped in 2026. Development happens publicly on GitHub with 15 commits in the last 90 days. Among its 5 catalogued features are LLM benchmarking, hardware advisor, and token/sec measurement.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do