PulseGateCategoriesMethodologyCompanyThe global software index— through the gate this hour
Coverage—in the index
Freshness—newest listing
Cadence—last week average · — today
Index9 markets90 categories · 139 niches
PulseGate

The global index of software taking shape now.

Stores show what passed through a store. Launch sites show what launched there. Catalogs show what entered their catalog. Each sees the market through its own gate. PulseGate reads across them.

FollowGitHubX (Twitter)LinkedIn
Platform
IndexIndex by setCategoriesIndustry UpdatesMethodologySupply IndexData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutTeamDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateLLllm-agent-bench
Visit↗
Skip to content
  1. Index›
  2. Autonomous agents & workflows›
  3. llm-agent-bench
← Back to the index
LL

llm-agent-bench

PyPI·Infrastructure

llm-agent-bench is an open-source CLI tool for benchmarking autonomous AI agents on task completion, tool use, goal adherence, and safety. It works with any agent by providing a callable interface, supporting AI researchers and developers.

Open SourceMITWebCLILinuxmacOSWindows
Visit PyPI↗

Overview

4 features

llm-agent-bench sits in PulseGate's Autonomous agents & workflows category. It focuses on evaluating and benchmarking the performance and safety of autonomous AI agents. It is built as an open-source project for AI researchers and developers. llm-agent-bench is open source under the MIT license. It runs on the web, the command line, Linux, macOS, and Windows.

llm-agent-bench first shipped in 2026. Key capabilities include agent benchmarking, task evaluation, and tool use analysis.

Summary written by a language model from the project’s public pages.

  • ✓Agent benchmarking
  • ✓Task evaluation
  • ✓Tool use analysis
  • ✓Safety checks
Tags
agent-benchmarkingai-evaluationautonomous-agents
AI capabilities
CodeStructured

Built with & integrations

Runs on
BrowserCLILinuxmacOSWindows

Trust & compliance

License
MIT
Verified signals
✓HTTPS✓Open Source✓Free tier

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed25 Jun · 18:34 UTC
    Listing verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about llm-agent-bench

What does llm-agent-bench do?
Llm-agent-bench focuses on evaluating and benchmarking the performance and safety of autonomous AI agents. It is catalogued under Autonomous agents & workflows on PulseGate.
Who should use llm-agent-bench?
llm-agent-bench is an open-source project built for AI researchers and developers.
Does llm-agent-bench have a free plan?
Yes — llm-agent-bench is open source under the MIT license and free to use.
What platforms does llm-agent-bench run on?
llm-agent-bench runs on the web, the command line, Linux, macOS, and Windows.
Is llm-agent-bench still active?
Unverified. llm-agent-bench has not been re-checked since it entered the index, so there is no finding either way — and only a positive finding would say otherwise.
What are alternatives to llm-agent-bench?
Similar projects tracked by PulseGate include litebench, agentbench-cli, and llm-parliament.litebenchagentbench-clillm-parliamentSee all 18 alternatives
When did llm-agent-bench launch?
llm-agent-bench first shipped in 2026.
Is llm-agent-bench open source?
Yes — llm-agent-bench is open source under the MIT license.

At a glance

Platforms
Cli
Languages
English
License
MIT
Built for
AI researchers and developers
Model
Open source
Solves
Evaluating and benchmarking the performance and safety of autonomous AI agents.

Registered as

PyPI
llm-agent-bench

Developer

Lindaoraegbunam

Index record

Identity confidence
Low · 64
Indexed
25 Jun 2026
Lifecycle
Alive
Last seen
25 Jun 2026
Identity audit (12)
Slug
llm-agent-bench-pypi-org
Lifecycle last checked
8 Aug 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
25 Jun 2026
Timeline basis
Indexed-at chronology. This listing's first-seen date was written by the catalog backfill, not observed here, so it is not treated as a sighting.
Name from
Written by a language model from the project's public pages.
Category from
Recorded source: taxonomy-pass-a.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model from page content.
Canonical URL
https://pypi.org/project/llm-agent-bench

Ship this? Send a correction — no account, and you get a link to follow it.

Similar projects

Closest matches by what these projects do

  • LIlitebenchgithub.com
  • AGagentbench-clipypi.org
  • LLllm-parliamentpypi.org
  • BEbenchflowgithub.com
  • LOlocal-bench-ailocal-bench.ai
  • LLllm-bench-costgithub.com

See 18 alternatives to llm-agent-bench →