caliper-eval
PulseGate's liveness check found it on 14 Sep 2026; it is registered on GitHub and PyPI and has been in the index since 3 Jul 2026. How this is checked
caliper-eval is an open-source CLI tool designed for evaluating Claude Code skills and AI agents. It provides developers and researchers with tools to assess and benchmark AI agent performance from the command line.
Inferred · not functionally tested
Overview
5 featuresPurpose: Enabling developers to evaluate and benchmark Claude Code skills and AI agents via the command line.
Inferred · not functionally tested
Audience: AI developers and researchers
Inferred · not functionally tested
Functions: agents
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: unknown
Recorded constraints: pricing: open_source · license: MIT · platforms: CLI · deployment: cli
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: pypi.org · github.com. These links do not verify the individual claims.
caliper-eval is an Agent evaluation & testing project. Inferred · not functionally tested: Enabling developers to evaluate and benchmark Claude Code skills and AI agents via the command line. Inferred · not functionally tested: caliper-eval is an open-source project aimed at AI developers and researchers. Basis unknown · not verified: The project is open source (MIT). Basis unknown · not verified: caliper-eval is available on the command line.
edonadei builds and maintains caliper-eval, and it first shipped in 2026. The project is developed in the open on GitHub with 23 stars and 108 commits in the last 90 days. Inferred · not functionally tested: Among its 5 catalogued features are AI evaluation, Claude Code skills, and agent testing.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- AI evaluation
- Claude Code skills
- Agent testing
- CLI interface
- Skill assessment
Topics: Inferred · not functionally tested
Built with & integrations
Trust & compliance
Indexing history
5What PulseGate has recorded for this listing
Frequently asked questions about caliper-eval
- What does caliper-eval do?
- Inferred · not functionally tested: Enabling developers to evaluate and benchmark Claude Code skills and AI agents via the command line. It is catalogued under Agent evaluation & testing on PulseGate.
- Who should use caliper-eval?
- Inferred · not functionally tested: caliper-eval is an open-source project built for AI developers and researchers.
- Does caliper-eval have a free plan?
- Basis unknown · not verified: Yes — caliper-eval is open source under the MIT license and free to use.
- What platforms does caliper-eval run on?
- Basis unknown · not verified: caliper-eval runs on the command line.
- Is caliper-eval still active?
- PulseGate's liveness check found it on 14 Sep 2026. Its GitHub repository shows 108 commits in the last 90 days.
- What projects are similar to caliper-eval?
- Similar projects tracked by PulseGate include Coder Eval, agent-skill-eval, and coder-eval.Coder Evalagent-skill-evalcoder-eval
- Who develops caliper-eval?
- caliper-eval is developed by edonadei.
- When did caliper-eval launch?
- caliper-eval first shipped in 2026.
Similar projects
Closest matches by what these projects do