Skip to content
Back to the index

skillrig

PyPIInfrastructure

PulseGate's liveness check found it on 3 Oct 2026; it is registered on PyPI and has been in the index since 19 Aug 2026. How this is checked

skillrig is a Python package and command-line tool for testing agent skills against real coding agents. It runs prompts, checks observable outcomes with assertions, and uses an LLM to grade results that require qualitative evaluation.

Inferred · not functionally tested

Open SourceMITCLISelf-hosted
Visit PyPI

Overview

6 features

Purpose: Evaluating whether agent skills work reliably across real coding-agent runs.

Inferred · not functionally tested

Audience: AI agent developers and evaluators

Inferred · not functionally tested

Functions: agents, code_generation

Inferred · not functionally tested

Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)

Recorded constraints: pricing: open_source · license: MIT · platforms: CLI · deployment: cli, self_hosted

Constraint provenance is unknown; confirm requirements with the publisher.

Record sources: pypi.org. These links do not verify the individual claims.

skillrig is an Agent evaluation & testing project. Inferred · not functionally tested: It focuses on evaluating whether agent skills work reliably across real coding-agent runs. Inferred · not functionally tested: skillrig is an open-source project aimed at AI agent developers and evaluators. Basis unknown · not verified: The project is open source (MIT). Basis unknown · not verified: It runs on the command line, and it can be self-hosted.

skillrig first shipped in 2026. Inferred · not functionally tested: Key capabilities include prompt execution, outcome assertions, and LLM grading.

Summary written by a language model from the project’s public pages.

Tasks: Inferred · not functionally tested

  • Prompt execution
  • Outcome assertions
  • LLM grading
  • Coding-agent testing
  • Pytest integration
  • Agent skill evaluation

Topics: Inferred · not functionally tested

Tags
agent-evaluationcoding-agentsskill-testingllm-grading
AI capabilities
TextCodeStructured
Inference: Cloud API

JSON profile · Text profile · Access guide

Built with & integrations

Runs on
CLISelf-hosted

Trust & compliance

License
MIT
Public signals
HTTPSOpen Source

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed19 Aug · 15:12 UTC
    skillrig seen via PyPI Fresh Feed
    Source: PyPI Fresh Feed · Open

Frequently asked questions about skillrig

What is skillrig?
Inferred · not functionally tested: Skillrig focuses on evaluating whether agent skills work reliably across real coding-agent runs. It is catalogued under Agent evaluation & testing on PulseGate.
Who is skillrig for?
Inferred · not functionally tested: skillrig is an open-source project built for AI agent developers and evaluators.
Is skillrig free?
Basis unknown · not verified: Yes — skillrig is open source under the MIT license and free to use.
What platforms does skillrig run on?
Basis unknown · not verified: skillrig runs on the command line. It can also be self-hosted.
Is skillrig still active?
PulseGate's liveness check found it on 3 Oct 2026.
What are alternatives to skillrig?
Similar projects tracked by PulseGate include skill-quality-lab, skill-lab, and pytest-skillcheck.skill-quality-labskill-labpytest-skillcheck
When did skillrig launch?
skillrig first shipped in 2026.
Is skillrig open source?
Basis unknown · not verified: Yes — skillrig is open source under the MIT license.

Similar projects

Closest matches by what these projects do