agentsec-eval is an open-source CLI framework for evaluating the security of AI agents. It provides adversarial test runners, server-side audits, and scoring mechanisms to help researchers and developers identify vulnerabilities such as prompt injection and improve agent robustness.
In the LLM eval & observability space, agentsec-eval takes a focused approach. It focuses on assessing and improving the security of AI agents through adversarial testing and audits. It is built as an open-source project for AI security researchers and developers. The project is open source (MIT). It runs on the command line, and it can be self-hosted.
It is developed by raoliaoyuan, and it first shipped in 2026. Key capabilities include adversarial testing, security audit, and scoring system.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do