harness-bench-fast provides a standardized, self-contained benchmark with 298 tasks covering file operations, code editing, CSV/SQLite/XLSX data pipelines, memory management, and other agentic scenarios. It is designed for rigorous evaluation of AI agents and frameworks such as LangChain. The package is open source under the MIT license and available on PyPI and GitHub.
harness-bench-fast is a LLM eval & observability product. It focuses on evaluating and comparing the performance of AI agents on realistic software engineering and data tasks. It is built as an open-source project for AI researchers and developers. harness-bench-fast is open source under the MIT license. It runs on the command line.
ai-forever builds and maintains harness-bench-fast, and the product first shipped in 2026. Development happens publicly on GitHub with 37 stars and 76 commits in the last 90 days. Key capabilities include Agent Benchmark, File Operations, and Code Editing.
Latest indexed changes and source events
harness-bench-fast verified by the PulseGate indexer
Other apps tracked under the same category.