nkama-fact-benchmark is an open-source benchmark tool designed to test whether AI assistants can provide verifiable evidence for their claims. It supports evidence gating, prompt engineering, and agent evaluation, making it useful for AI researchers and developers focused on model reliability and trustworthiness.
In the LLM eval & observability space, nkama-fact-benchmark takes a focused approach. It focuses on evaluating whether AI assistants can provide verifiable evidence for their claims. It is built as an open-source project for AI researchers and developers. nkama-fact-benchmark is open source under the Apache-2.0 license. The product ships for the command line.
It is developed by donkk11, and the product first shipped in 2026. Development happens publicly on GitHub with 14 commits in the last 90 days. Key capabilities include evidence gating, claim verification, and prompt engineering.
Latest indexed changes and source events
nkama-fact-benchmark verified by the PulseGate indexer
Other apps tracked under the same category.