Skip to content
Alternatives
Software like agentanvil
What else does this job. Matched on what each project does, not on who links to whom.
Closest first
- agentlint-devpypi.orgLint, score, and optimize AI agent configuration files (CLAUDE.md, AGENTS.md, Cursor rules)
- agent-beltgithub.comEvaluation harness for real headless CLI agents - reproducible multi-turn scenarios, rule + LLM scoring, cross-agent comparison
- agent-evalpypi.orgAgent evaluation toolkit
- agentlintpypi.orgagentlint is an open-source CLI tool that provides real-time quality guardrails and linting for AI coding agents. It helps developers maintain code quality and safety by integrating with agentic coding workflows and enforcing best practices.
- agentkit-cligithub.comUnified CLI for the Agent Quality Toolkit (agentmd, coderace, agentlint, agentreflect)
- playagentgithub.comAgent testing SDK. Instrument, assert, classify.
- agentbreakgithub.comChaos proxy for testing LLM agent resilience — inject faults like latency, errors, and bad responses. Supports OpenAI, Anthropic, and MCP.
- agentbeacongithub.comMulti-agent orchestrator for AI coding tools
- AgentXgithub.comAgentX - Native compiled Python agent framework
- agentchantigithub.comA multi-agent AI coding system with built-in RAG — local or cloud, CLI or API.
- uAgentsgithub.comLightweight framework for rapid agent-based development
- agentycgithub.comDeterministic MCP-first browser automation runtime for coding agents
- nnn-agentgithub.comMulti-agent coding system powered by local LLMs
- agentircgithub.comAI-powered IRC agent with multi-provider LLM support
- agent-clipypi.orgA suite of AI-powered command-line tools for text correction, audio transcription, and voice assistance.
- DefenseAgentgithub.comMulti-LLM agent framework with mem0-backed memory, llama-index RAG, MCP tool support, and reflection.
- agent-interrogatorpypi.orgAn AI agent interrogation framework for identifying attack surface.
- openagentdgithub.comSelf-hosted AI agents — streaming chat, tool use, persistent memory, multi-agent teams
- agent-minigithub.comagent-mini is an open-source CLI tool that acts as a lightweight personal AI agent, automating shell, file, and web tasks. It supports both local and cloud LLMs, making it ideal for developers and power users seeking AI-driven automation from the command line.
- dotagent-toolkitpypi.orgCLI-Toolkit für AI-Agent-Infrastruktur: .agent-Ordner, Tool-Discovery, Credential-Management
- agentpoolpypi.orgPydantic-AI based Multi-Agent Framework with YAML-based Agents, Teams, Workflows & Extended ACP / AGUI integration
- agent-failure-debuggerpypi.orgDiagnose why your LLM agent failed. Deterministic causal analysis with fix generation.
- polyagent-aigithub.compolyagent-ai is an open-source multi-agent framework that automates code generation, API testing, and UI automation tasks. It leverages LLMs and agent-based workflows, making it suitable for developers and QA engineers seeking to streamline automation. Licensed under MIT.
- Phantom Agentgithub.comAutonomous Offensive Security Intelligence - AI-powered penetration testing
- bareagent-clipypi.orgbareagent-cli is a pure Python command-line code agent that supports multiple LLM providers, fine-grained permission systems, multi-agent coordination, and an extensible skill architecture. It allows developers to run autonomous coding tasks in the terminal while maintaining control over permissions and agent interactions. Fully open source under MIT license.
- AgentEvalagenteval.devAI. NET ecosystem. The platform provides features such as tool usage validation, which allows users to assert on tool chains and verify that specific tools are called in the correct order with appropriate arguments. Stochastic evaluation is supported, enabling repeated runs of agent tasks to assess actual success rates and standard deviations, reflecting the non-deterministic nature of large language models. Workflow evaluation capabilities allow for the testing of multi-agent flows, including validation of executor order, edge traversal, and per-graph tool calls. Performance evaluation tools enable users to set and assert on service level agreements (SLAs) related to response times, total duration, and estimated costs. AgentEval includes model comparison functionality, letting users benchmark multiple models against defined metrics such as tool accuracy, relevance, and cost per request. The toolkit also supports recording and replaying agent interactions, which allows for consistent, repeatable evaluations without incurring additional API costs. Security evaluation is addressed through a Red Team module that tests agents against 258 attack probes across all 10 OWASP LLM Top 10 vulnerabilities, with MITRE ATLAS technique mapping. This module covers a wide range of attack types, including prompt injection, jailbreaks, PII leakage, and more, and supports both quick scans and advanced, customizable attack pipelines. Security compliance reports can be exported in PDF format. Memory evaluation is another key feature, with tools for benchmarking agent memory retention, recall depth, temporal reasoning, fact updates, cross-session persistence, and noise resistance. Results can be exported as interactive HTML reports. NET developers seeking to rigorously test, benchmark, and ensure the reliability, security, and performance of their AI agents before production use.
- janus-labsgithub.com3DMark for AI Agents - Profile AI coding agent capabilities across code quality, error resilience, and instruction resilience
- agentscore-clipypi.orgLighthouse for AI agent development environments
- agentos-frameworkpypi.orgagentos-framework is an open-source agent framework for developing, orchestrating, and managing autonomous AI agents. It supports multiple LLM providers, function calling, streaming, checkpointing, and swarm coordination, making it suitable for advanced AI development and research.
- agentype-cligithub.comLocal AI-agent usage analytics and persona archetypes
- agentcfggithub.comCLI tool for deploying and managing AI coding agent configurations (MCP servers, skills, instructions) across multiple providers.
- all-in-agentsgithub.comA minimal, universal agent framework. Zero mandatory dependencies.
- agentfluentgithub.comLocal-first agent analytics with prompt diagnostics
- agent-notesgithub.comAI agent configuration manager for Claude Code, OpenCode, and Copilot
- agent-zoogithub.comSecurity harness for AI coding agents (Claude Code, Codex CLI, etc.) — mitmproxy payload inspection + TOML policy control.
- agent-tool-routergithub.comPick the right tools for an agent task. Boring baseline. Open dataset.
- agent-def-translatorgithub.comTranslate canonical coding-agent definitions into platform-native agent artifacts.
Ranked by how close each one sits to agentanvil in the index, not by popularity. Back to agentanvil →