Skip to content
Back to the index

aau-harness

PyPIInfrastructure

PulseGate's liveness check found it on 3 Oct 2026; it is registered on PyPI and has been in the index since 23 Aug 2026. How this is checked

aau-harness is an Apache-2.0 Python package for provider-neutral evaluation of AI agents. It supports seeded scenarios, exact scoring, bring-your-own-agent adapters, repeated runs, cost tracking, and evaluation receipts for developers and researchers.

Inferred · not functionally tested

Open SourceApache-2.0CLISelf-hosted
Visit PyPI

Overview

6 features

Purpose: Evaluating AI agents consistently across seeded scenarios, repeated runs, costs, and receipts.

Inferred · not functionally tested

Audience: AI agent developers and evaluators

Inferred · not functionally tested

Functions: analytics

Inferred · not functionally tested

Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)

Recorded constraints: pricing: open_source · license: Apache-2.0 · platforms: CLI · deployment: cli, self_hosted

Constraint provenance is unknown; confirm requirements with the publisher.

Record sources: pypi.org. These links do not verify the individual claims.

In the Agent evaluation & testing space, aau-harness takes a focused approach. Inferred · not functionally tested: It focuses on evaluating AI agents consistently across seeded scenarios, repeated runs, costs, and receipts. Inferred · not functionally tested: It is built as an open-source project for AI agent developers and evaluators. Basis unknown · not verified: The project is open source (Apache-2.0). Basis unknown · not verified: It ships for the command line, and it can be self-hosted.

aau-harness first shipped in 2026. Inferred · not functionally tested: Key capabilities include seeded scenarios, exact scoring, and agent adapters.

Summary written by a language model from the project’s public pages.

Tasks: Inferred · not functionally tested

  • Seeded scenarios
  • Exact scoring
  • Agent adapters
  • Repeated runs
  • Cost tracking
  • Evaluation receipts

Topics: Inferred · not functionally tested

Tags
agent-evaluationllm-testingscenario-benchmarkingagent-adapters
AI capabilities
TextStructured

JSON profile · Text profile · Access guide

Built with & integrations

Runs on
CLISelf-hosted

Trust & compliance

License
Apache-2.0
Public signals
HTTPSOpen SourceFree tier

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed23 Aug · 21:06 UTC
    aau-harness seen via PyPI Fresh Feed
    Source: PyPI Fresh Feed · Open

Frequently asked questions about aau-harness

What is aau-harness?
Inferred · not functionally tested: Aau-harness focuses on evaluating AI agents consistently across seeded scenarios, repeated runs, costs, and receipts. It is catalogued under Agent evaluation & testing on PulseGate.
Who should use aau-harness?
Inferred · not functionally tested: aau-harness is an open-source project built for AI agent developers and evaluators.
Is aau-harness free?
Basis unknown · not verified: Yes — aau-harness is open source under the Apache-2.0 license and free to use.
What platforms does aau-harness run on?
Basis unknown · not verified: aau-harness runs on the command line. It can also be self-hosted.
Is aau-harness still active?
PulseGate's liveness check found it on 3 Oct 2026.
What projects are similar to aau-harness?
Similar projects tracked by PulseGate include aehf, ai-harness-cli, and agent-security-harness.aehfai-harness-cliagent-security-harness
How long has aau-harness been around?
aau-harness first shipped in 2026.
Is aau-harness open source?
Basis unknown · not verified: Yes — aau-harness is open source under the Apache-2.0 license.

Similar projects

Closest matches by what these projects do