Evals Coach
PulseGate's liveness check found it on 8 Oct 2026; it is registered on GitHub and has been in the index since 7 Sep 2026. How this is checked
Evals Coach is a Claude plugin that helps AI product managers turn feature descriptions or real outputs into runnable evaluations. It guides users through criteria, test cases, graders, judge prompts, and release gates, producing outputs that can be used with existing evaluation stacks.
Inferred · not functionally tested
Overview
6 featuresPurpose: Designing reliable AI evaluations without needing specialized eval engineering expertise.
Inferred · not functionally tested
Audience: AI product managers
Inferred · not functionally tested
Functions: analytics
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: unknown
Recorded constraints: pricing: unknown · license: Proprietary · platforms: CLI, WEB · deployment: browser, cli, cloud_managed
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: evalscoach.com · github.com. These links do not verify the individual claims.
In the Agent evaluation & testing space, Evals Coach takes a focused approach. Inferred · not functionally tested: It focuses on designing reliable AI evaluations without needing specialized eval engineering expertise. Inferred · not functionally tested: It is built as a B2B product for AI product managers. Basis unknown · not verified: It runs on the web and the command line.
Evals Coach first shipped in 2026. The project is developed in the open on GitHub with 82 commits in the last 90 days. Inferred · not functionally tested: Among its 6 catalogued features are feature description intake, evaluation questions, and test case design.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- Feature description intake
- Evaluation questions
- Test case design
- Failure mode tracking
- Grader design
- Judge prompts
Topics: Inferred · not functionally tested
Built with & integrations
- Claude Code
- commit 7448a92304a6 · since Sep 2026
- openai
- bgpt- in the HTML
- Vercel
- x-vercel-id header · x-vercel-cache header
Trust & compliance
Indexing history
1What PulseGate has recorded for this listing
- Indexed7 Sep · 20:14 UTCShow HN: Evals Coach – a Claude plugin to help PMs write good evals seen via Hacker News firehose (Algolia)Source: Hacker News firehose (Algolia) · Open
Frequently asked questions about Evals Coach
- What is Evals Coach?
- Inferred · not functionally tested: Evals Coach focuses on designing reliable AI evaluations without needing specialized eval engineering expertise. It is catalogued under Agent evaluation & testing on PulseGate.
- Who should use Evals Coach?
- Inferred · not functionally tested: Evals Coach is a B2B product built for AI product managers.
- What platforms does Evals Coach run on?
- Basis unknown · not verified: Evals Coach runs on the web and the command line.
- Is Evals Coach still maintained?
- PulseGate's liveness check found it on 8 Oct 2026. Its GitHub repository shows 82 commits in the last 90 days.
- When did Evals Coach launch?
- Evals Coach first shipped in 2026.
- Is Evals Coach open source?
- Basis unknown · not verified: Evals Coach has a public GitHub repository.
Similar projects
Closest matches by what these projects do