litebench
PulseGate's liveness check found it on 13 Sep 2026; it is registered on GitHub and PyPI and has been in the index since 27 Jun 2026. How this is checked
litebench is an open-source CLI tool that enables developers and researchers to benchmark large language models and AI agents. It supports quick setup and evaluation workflows, including popular benchmarks like GSM8K and HumanEval.
Inferred · not functionally tested
Overview
5 featuresPurpose: Providing a fast and easy way to benchmark and evaluate LLMs and AI agents for developers and researchers.
Inferred · not functionally tested
Audience: AI researchers and developers
Inferred · not functionally tested
Functions: analytics
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: unknown
Recorded constraints: pricing: open_source · license: MIT · platforms: CLI · deployment: cli
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: github.com. These links do not verify the individual claims.
litebench is a LLM evaluation & benchmarks project. Inferred · not functionally tested: It focuses on providing a fast and easy way to benchmark and evaluate LLMs and AI agents for developers and researchers. Inferred · not functionally tested: litebench is an open-source project aimed at AI researchers and developers. Basis unknown · not verified: The project is open source (MIT). Basis unknown · not verified: It runs on the command line.
It is developed by he-yufeng, and it first shipped in 2026. The project is developed in the open on GitHub with 14 commits in the last 90 days. Inferred · not functionally tested: Among its 5 catalogued features are benchmark runner, LLM evaluation, and agent evaluation.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- Benchmark runner
- LLM evaluation
- Agent evaluation
- Quick setup
- Supports GSM8K and HumanEval
Topics: Inferred · not functionally tested
Built with & integrations
Trust & compliance
Indexing history
What PulseGate has recorded for this listing
Frequently asked questions about litebench
- What does litebench do?
- Inferred · not functionally tested: Litebench focuses on providing a fast and easy way to benchmark and evaluate LLMs and AI agents for developers and researchers. It is catalogued under LLM evaluation & benchmarks on PulseGate.
- Who is litebench for?
- Inferred · not functionally tested: litebench is an open-source project built for AI researchers and developers.
- Does litebench have a free plan?
- Basis unknown · not verified: Yes — litebench is open source under the MIT license and free to use.
- What platforms does litebench run on?
- Basis unknown · not verified: litebench runs on the command line.
- Is litebench still active?
- PulseGate's liveness check found it on 13 Sep 2026. Its GitHub repository shows 14 commits in the last 90 days.
- What projects are similar to litebench?
- Similar projects tracked by PulseGate include llm-agent-bench, bench-my-llm, and labbench-cli.llm-agent-benchbench-my-llmlabbench-cli
- Who makes litebench?
- litebench is developed by he-yufeng.
- How long has litebench been around?
- litebench first shipped in 2026.
Similar projects
Closest matches by what these projects do