agent-eval-flow
PulseGate's liveness check found it on 15 Sep 2026; it is registered on GitHub and PyPI and has been in the index since 15 Sep 2026. How this is checked
agent-eval-flow is a Python package for evaluating complete agent systems while preserving evidence from their executions. It is intended for developers testing and analyzing LLM-powered agents and their workflows.
Inferred · not functionally tested
Overview
4 featuresPurpose: Evaluating complete agent systems while preserving execution evidence for reproducible analysis.
Inferred · not functionally tested
Audience: AI and agent developers
Inferred · not functionally tested
Functions: agents, analytics
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)
Recorded constraints: pricing: open_source · license: Open Source · platforms: CLI · deployment: cli, self_hosted
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: pypi.org · github.com. These links do not verify the individual claims.
In the Agent evaluation & testing space, agent-eval-flow takes a focused approach. Inferred · not functionally tested: It focuses on evaluating complete agent systems while preserving execution evidence for reproducible analysis. Inferred · not functionally tested: It is built as an open-source project for AI and agent developers. Basis unknown · not verified: The project is open source (Open Source). Basis unknown · not verified: It runs on the command line, and it can be self-hosted.
Guy Bass builds and maintains agent-eval-flow, and it first shipped in 2026. The project is developed in the open on GitHub with 6 commits in the last 90 days. Inferred · not functionally tested: Among its 4 catalogued features are agent evaluation, execution evidence, and LLM evaluation.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- Agent evaluation
- Execution evidence
- LLM evaluation
- Agent workflow analysis
Topics: Inferred · not functionally tested
Built with & integrations
Trust & compliance
Indexing history
2What PulseGate has recorded for this listing
- Indexed15 Sep · 12:21 UTCagent-eval-flow seen via PyPI Bulk EnumeratorSource: PyPI Bulk Enumerator · Open
Frequently asked questions about agent-eval-flow
- What is agent-eval-flow?
- Inferred · not functionally tested: Agent-eval-flow focuses on evaluating complete agent systems while preserving execution evidence for reproducible analysis. It is catalogued under Agent evaluation & testing on PulseGate.
- Who should use agent-eval-flow?
- Inferred · not functionally tested: agent-eval-flow is an open-source project built for AI and agent developers.
- Is agent-eval-flow free?
- Basis unknown · not verified: Yes — agent-eval-flow is open source under the Open Source license and free to use.
- What platforms does agent-eval-flow run on?
- Basis unknown · not verified: agent-eval-flow runs on the command line. It can also be self-hosted.
- Is agent-eval-flow still active?
- PulseGate's liveness check found it on 15 Sep 2026. Its GitHub repository shows 6 commits in the last 90 days.
- What are alternatives to agent-eval-flow?
- Similar projects tracked by PulseGate include evidenceflow, agentaudit-eval, and agent-evaluation-lab.evidenceflowagentaudit-evalagent-evaluation-lab
- Who makes agent-eval-flow?
- agent-eval-flow is developed by Guy Bass.
- When did agent-eval-flow launch?
- agent-eval-flow first shipped in 2026.
Similar projects
Closest matches by what these projects do