agentgrade is an open-source testing framework for multi-agent AI systems. It provides regression tests, credit assignment, and prompt patching, enabling developers and researchers to evaluate and improve agent workflows efficiently.
In the AI space, agentgrade takes a focused approach. It focuses on testing and evaluating multi-agent AI workflows with automated regression and prompt patching. agentgrade is an open-source project aimed at AI developers and researchers. The project is open source (MIT). agentgrade is available on the command line, and it can be self-hosted.
Behind agentgrade is Shengyong Niu, and it first shipped in 2026. Key capabilities include regression testing, prompt patching, and credit assignment.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do