PulseGateCategoriesMethodologyCompanyThe global software index— through the gate this hour
Coverage—in the index
Freshness—newest listing
Cadence—last week average · — today
Index9 markets90 categories · 139 niches
PulseGate

The global index of software taking shape now.

Stores show what passed through a store. Launch sites show what launched there. Catalogs show what entered their catalog. Each sees the market through its own gate. PulseGate reads across them.

FollowGitHubX (Twitter)LinkedIn
Platform
IndexIndex by setCategoriesIndustry UpdatesMethodologySupply IndexData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutTeamDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateAGagent-skill-eval
Visit↗
Skip to content
  1. Index›
  2. LLM eval & observability›
  3. agent-skill-eval
← Back to the index
AG

agent-skill-eval

PyPI·Infrastructure

agent-skill-eval is an open-source CLI framework for evaluating the skills of code-generating agents across models like OpenCode, Claude Code, and Codex. It enables researchers and developers to benchmark agent performance using standardized tests.

Open SourceMITCLI
Visit PyPI↗

Overview

4 features

In the LLM eval & observability space, agent-skill-eval takes a focused approach. It focuses on providing a standardized way to evaluate and benchmark agent skills across different LLM code models. agent-skill-eval is an open-source project aimed at AI researchers and developers. agent-skill-eval is open source under the MIT license. agent-skill-eval is available on the command line.

Behind agent-skill-eval is tardigrde, and it first shipped in 2026. Development happens publicly on GitHub with 53 commits in the last 90 days. Key capabilities include agent benchmarking, skill evaluation, and LLM support.

Summary written by a language model from the project’s public pages.

  • ✓Agent benchmarking
  • ✓Skill evaluation
  • ✓LLM support
  • ✓Command-line interface
Tags
agent-evaluationllm-benchmarkingcode-eval
AI capabilities
Code

Built with & integrations

AI providers
openai
Runs on
CLI

Trust & compliance

License
MIT
Verified signals
✓HTTPS✓Open Source✓Free tier✓GitHub✓Active maintenance

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed30 Jun · 08:53 UTC
    agent-skill-eval verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about agent-skill-eval

What does agent-skill-eval do?
Agent-skill-eval focuses on providing a standardized way to evaluate and benchmark agent skills across different LLM code models. It is catalogued under LLM eval & observability on PulseGate.
Who is agent-skill-eval for?
agent-skill-eval is an open-source project built for AI researchers and developers.
Is agent-skill-eval free?
Yes — agent-skill-eval is open source under the MIT license and free to use.
What platforms does agent-skill-eval run on?
agent-skill-eval runs on the command line.
Is agent-skill-eval still active?
The GitHub repository shows 53 commits in the last 90 days.
What are alternatives to agent-skill-eval?
Similar projects tracked by PulseGate include agent-skill-description-optimizer, openagent-eval, and OpenAgentSkill.agent-skill-description-optimizeropenagent-evalOpenAgentSkill
Who makes agent-skill-eval?
agent-skill-eval is developed by tardigrde.
How long has agent-skill-eval been around?
agent-skill-eval first shipped in 2026.

At a glance

Platforms
Cli
Languages
English
Open source
Yes (GitHub)
License
MIT
Built for
AI researchers and developers
Model
Open source
Solves
Providing a standardized way to evaluate and benchmark agent skills across different LLM code models.

Registered as

GitHub
tardigrde/agent-skill-eval
PyPI
agent-skill-eval

Developer

tardigrde
Small team
↗ GitHub

Open source

View on GitHub →
Stars
0
Forks
0
Open issues
0
Last commit
12 Jun 2026
Commits 90d
53
Contributors
2
Authorship
Small team
Default branch
main
Latest release
v0.6.1 · 12 Jun 2026

Index record

Identity confidence
Low · 64
Indexed
30 Jun 2026
Lifecycle
Alive
Last seen
30 Jun 2026
Identity audit (12)
Slug
agent-skill-eval-pypi-org
Lifecycle last checked
8 Aug 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
30 Jun 2026
Timeline basis
Indexed-at chronology. This listing's first-seen date was written by the catalog backfill, not observed here, so it is not treated as a sighting.
Name from
Written by a language model from the project's public pages.
Category from
Assigned by a language model.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model from page content.
Canonical URL
https://pypi.org/project/agent-skill-eval

Ship this? Send a correction — no account, and you get a link to follow it.

Similar projects

Closest matches by what these projects do

  • AGagent-skill-description-optimizerpypi.org
  • OPopenagent-evalpypi.org
  • OPOpenAgentSkillopenagentskill.com
  • AGagent-skill-installergithub.com
  • SKskill-labgithub.com
  • CAcaliper-evalpypi.org