litebench is an open-source CLI tool that enables developers and researchers to benchmark large language models and AI agents. It supports quick setup and evaluation workflows, including popular benchmarks like GSM8K and HumanEval.
litebench is an AI project. It focuses on providing a fast and easy way to benchmark and evaluate LLMs and AI agents for developers and researchers. litebench is an open-source project aimed at AI researchers and developers. The project is open source (MIT). It runs on the command line.
It is developed by he-yufeng, and it first shipped in 2026. The project is developed in the open on GitHub with 14 commits in the last 90 days. Among its 5 catalogued features are benchmark runner, LLM evaluation, and agent evaluation.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do