agent-exam
PulseGate's liveness check found it on 1 Oct 2026; it is registered on GitHub and PyPI and has been in the index since 12 Aug 2026. How this is checked
agent-exam is an Apache-2.0 evaluation framework for testing agent skills across Claude Code, Codex CLI, Copilot CLI, and OpenCode. It is intended for developers building and benchmarking agent workflows.
Inferred · not functionally tested
Overview
5 featuresPurpose: Evaluating and comparing agent skills consistently across multiple coding-agent environments.
Inferred · not functionally tested
Audience: developers building and evaluating AI agent workflows
Inferred · not functionally tested
Functions: analytics
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)
Recorded constraints: pricing: open_source · license: Apache-2.0 · platforms: CLI, WEB · deployment: browser, cli, self_hosted
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: agent-exam.readthedocs.io · github.com. These links do not verify the individual claims.
In the Agent evaluation & testing space, agent-exam takes a focused approach. Inferred · not functionally tested: It focuses on evaluating and comparing agent skills consistently across multiple coding-agent environments. Inferred · not functionally tested: It is built as an open-source project for developers building and evaluating AI agent workflows. Basis unknown · not verified: The project is open source (Apache-2.0). Basis unknown · not verified: agent-exam is available on the web and the command line, and it can be self-hosted.
It is developed by Zyte Data, and it first shipped in 2026. Development happens publicly on GitHub with 10 commits in the last 90 days. Inferred · not functionally tested: Among its 5 catalogued features are agent skill evaluation, cross-agent testing, and CLI evaluation.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- Agent skill evaluation
- Cross-agent testing
- CLI evaluation
- Benchmarking
- Evaluation reports
Topics: Inferred · not functionally tested
Built with & integrations
Trust & compliance
Indexing history
1What PulseGate has recorded for this listing
- Indexed12 Aug · 21:28 UTCagent-exam seen via PyPI Bulk EnumeratorSource: PyPI Bulk Enumerator · Open
Frequently asked questions about agent-exam
- What does agent-exam do?
- Inferred · not functionally tested: Agent-exam focuses on evaluating and comparing agent skills consistently across multiple coding-agent environments. It is catalogued under Agent evaluation & testing on PulseGate.
- Who should use agent-exam?
- Inferred · not functionally tested: agent-exam is an open-source project built for developers building and evaluating AI agent workflows.
- Is agent-exam free?
- Basis unknown · not verified: Yes — agent-exam is open source under the Apache-2.0 license and free to use.
- What platforms does agent-exam run on?
- Basis unknown · not verified: agent-exam runs on the web and the command line. It can also be self-hosted.
- Is agent-exam still active?
- PulseGate's liveness check found it on 1 Oct 2026. Its GitHub repository shows 10 commits in the last 90 days.
- What are alternatives to agent-exam?
- Similar projects tracked by PulseGate include agent-skill-eval, agent-evaluation-lab, and agent-code.agent-skill-evalagent-evaluation-labagent-code
- Who makes agent-exam?
- agent-exam is developed by Zyte Data.
- How long has agent-exam been around?
- agent-exam first shipped in 2026.
Similar projects
Closest matches by what these projects do