Skip to content
Back to the index

SagaBench

sagabench.comInfrastructure

PulseGate's liveness check found it on 3 Oct 2026; it is registered on PyPI and has been in the index since 14 Aug 2026. How this is checked

SagaBench provides client and reporting tools for evaluating AI agents over extended periods while they operate autonomously. It is intended for developers and researchers measuring agent behavior, outcomes, and reliability in long-horizon tasks.

Inferred · not functionally tested

Open SourceApache-2.0WebCLISelf-hosted
SagaBench preview
Visit sagabench.com

Overview

4 features

Purpose: Measuring how AI agents behave and perform when left in charge of tasks over time.

Inferred · not functionally tested

Audience: AI developers and researchers

Inferred · not functionally tested

Functions: analytics, monitoring

Inferred · not functionally tested

Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)

Recorded constraints: pricing: open_source · license: Apache-2.0 · platforms: CLI, WEB · deployment: browser, cli, self_hosted

Constraint provenance is unknown; confirm requirements with the publisher.

Record sources: sagabench.com. These links do not verify the individual claims.

SagaBench sits in PulseGate's Agent evaluation & testing category. Inferred · not functionally tested: It focuses on measuring how AI agents behave and perform when left in charge of tasks over time. Inferred · not functionally tested: SagaBench is an open-source project aimed at AI developers and researchers. Basis unknown · not verified: The project is open source (Apache-2.0). Basis unknown · not verified: It runs on the web and the command line, and it can be self-hosted.

SagaBench first shipped in 2026. Inferred · not functionally tested: Key capabilities include agent benchmarking, long-horizon evaluation, and measurement reports.

Summary written by a language model from the project’s public pages.

Tasks: Inferred · not functionally tested

  • Agent benchmarking
  • Long-horizon evaluation
  • Measurement reports
  • Client tools

Topics: Inferred · not functionally tested

Tags
agent-benchmarkinglong-horizon-evaluationagent-observabilityai-research

JSON profile · Text profile · Access guide

Built with & integrations

Runs on
BrowserCLISelf-hosted

Trust & compliance

License
Apache-2.0
Public signals
HTTPSOpen SourceFree tier

Indexing history

3

What PulseGate has recorded for this listing

  1. Indexed27 Sep · 11:42 UTC
    sagabench seen via PyPI Fresh Feed
    Source: PyPI Fresh Feed · Open
  2. Indexed25 Sep · 15:46 UTC
    sagabench seen via PyPI Fresh Feed
    Source: PyPI Fresh Feed · Open
  3. Indexed14 Aug · 19:25 UTC
    sagabench seen via PyPI Bulk Enumerator
    Source: PyPI Bulk Enumerator · Open

Frequently asked questions about SagaBench

What is SagaBench?
Inferred · not functionally tested: SagaBench focuses on measuring how AI agents behave and perform when left in charge of tasks over time. It is catalogued under Agent evaluation & testing on PulseGate.
Who is SagaBench for?
Inferred · not functionally tested: SagaBench is an open-source project built for AI developers and researchers.
Does SagaBench have a free plan?
Basis unknown · not verified: Yes — SagaBench is open source under the Apache-2.0 license and free to use.
What platforms does SagaBench run on?
Basis unknown · not verified: SagaBench runs on the web and the command line. It can also be self-hosted.
Is SagaBench still active?
PulseGate's liveness check found it on 3 Oct 2026.
What are alternatives to SagaBench?
Similar projects tracked by PulseGate include benchspec, CatchBench, and agentbench-cli.benchspecCatchBenchagentbench-cli
How long has SagaBench been around?
SagaBench first shipped in 2026.
Is SagaBench open source?
Basis unknown · not verified: Yes — SagaBench is open source under the Apache-2.0 license.

Similar projects

Closest matches by what these projects do