tourney is an open-source CLI tool for local-first benchmarking and evaluation of AI models. It allows users to run custom prompts against multiple models and compare their performance, making it ideal for AI researchers and developers focused on model evaluation and selection.
tourney is a LLM eval & observability project. It focuses on evaluating and benchmarking AI models locally using custom prompts to compare performance. It is built as an open-source project for AI researchers and developers needing model evaluation tools. tourney is open source under the Apache-2.0 license. It ships for the command line, and it can be self-hosted.
tourney first shipped in 2026. Key capabilities include benchmarking runner, prompt evaluation, and local execution.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Same category — not a similarity match