Aeval
PulseGate's liveness check found it on 3 Oct 2026; it is registered on GitHub and PyPI and has been in the index since 30 Aug 2026. How this is checked
Aeval is an MIT-licensed open-source framework for evaluating AI agents from OpenTelemetry traces. It provides evaluation suites, graders, metrics, storage, an API, and a command-line interface for developers testing agent systems.
Inferred · not functionally tested
Overview
6 featuresPurpose: Evaluating AI agent behavior consistently using traced runs, test suites, graders, and metrics.
Inferred · not functionally tested
Audience: AI developers and ML engineers building agent systems
Inferred · not functionally tested
Functions: analytics, monitoring
Inferred · not functionally tested
Interfaces: API: indicated (inferred, not tested) · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)
Recorded constraints: pricing: open_source · license: MIT · platforms: CLI · deployment: cli, self_hosted, api_only
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: pypi.org · github.com. These links do not verify the individual claims.
Aeval sits in PulseGate's Agent evaluation & testing category. Inferred · not functionally tested: It focuses on evaluating AI agent behavior consistently using traced runs, test suites, graders, and metrics. Inferred · not functionally tested: Aeval is an open-source project aimed at AI developers and ML engineers building agent systems. Basis unknown · not verified: Aeval is open source under the MIT license. Basis unknown · not verified: It runs on the command line and API, and it can be self-hosted.
Behind Aeval is QiYuyyds, and it first shipped in 2026. Development happens publicly on GitHub with 5 commits in the last 90 days. Inferred · not functionally tested: Key capabilities include OTel trace evaluation, evaluation suites, and custom graders. Inferred · not functionally tested: Catalogued interfaces include a public API.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- OTel trace evaluation
- Evaluation suites
- Custom graders
- Metrics
- Trace storage
- Evaluation API
Topics: Inferred · not functionally tested
Built with & integrations
Trust & compliance
Indexing history
2What PulseGate has recorded for this listing
- Indexed30 Aug · 08:48 UTCaeval-framework seen via PyPI Bulk EnumeratorSource: PyPI Bulk Enumerator · Open
Frequently asked questions about Aeval
- What does Aeval do?
- Inferred · not functionally tested: Aeval focuses on evaluating AI agent behavior consistently using traced runs, test suites, graders, and metrics. It is catalogued under Agent evaluation & testing on PulseGate.
- Who is Aeval for?
- Inferred · not functionally tested: Aeval is an open-source project built for AI developers and ML engineers building agent systems.
- Does Aeval have a free plan?
- Basis unknown · not verified: Yes — Aeval is open source under the MIT license and free to use.
- What platforms does Aeval run on?
- Basis unknown · not verified: Aeval runs on the command line and API. It can also be self-hosted.
- Is Aeval still active?
- PulseGate's liveness check found it on 3 Oct 2026. Its GitHub repository shows 5 commits in the last 90 days.
- What are alternatives to Aeval?
- Similar projects tracked by PulseGate include AgentEvals, evalite, and EvalView.AgentEvalsevaliteEvalView
- Who makes Aeval?
- Aeval is developed by QiYuyyds.
- When did Aeval launch?
- Aeval first shipped in 2026.
Similar projects
Closest matches by what these projects do