mcp-llm-eval
PulseGate's liveness check found it on 13 Sep 2026; it is registered on GitHub and PyPI and has been in the index since 29 Jun 2026. How this is checked
mcp-llm-eval is an open-source CLI tool and MCP server that packages LLM evaluation gates as reusable primitives for CI/CD workflows. It enables developers and ML engineers to automate and standardize the evaluation of large language models within their development pipelines. The tool supports benchmarking and integration with MCP for scalable model assessment.
Inferred · not functionally tested
Overview
5 featuresPurpose: Automating and standardizing LLM evaluation in CI/CD pipelines for developers and ML engineers.
Inferred · not functionally tested
Audience: machine learning engineers
Inferred · not functionally tested
Functions: monitoring, workflow_automation
Inferred · not functionally tested
Interfaces: API: indicated (inferred, not tested) · MCP: indicated (inferred, not tested) · CLI: indicated (inferred, not tested) · Self-hosting: indicated (inferred, not tested)
Recorded constraints: pricing: open_source · license: MIT · platforms: CLI · deployment: cli, api_only, self_hosted
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: pypi.org · github.com. These links do not verify the individual claims.
mcp-llm-eval sits in PulseGate's LLM evaluation & benchmarks category. Inferred · not functionally tested: It focuses on automating and standardizing LLM evaluation in CI/CD pipelines for developers and ML engineers. Inferred · not functionally tested: It is built as an open-source project for machine learning engineers. Basis unknown · not verified: The project is open source (MIT). Basis unknown · not verified: mcp-llm-eval is available on the command line and API, and it can be self-hosted.
It is developed by berkayildi, and it first shipped in 2026. The project is developed in the open on GitHub with 57 commits in the last 90 days. Inferred · not functionally tested: Key capabilities include LLM evaluation, CI/CD integration, and reusable gates. Inferred · not functionally tested: Catalogued interfaces include an MCP server and a public API.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- LLM evaluation
- CI/CD integration
- Reusable gates
- MCP server
- Benchmarking
Topics: Inferred · not functionally tested
Built with & integrations
Trust & compliance
Indexing history
What PulseGate has recorded for this listing
Frequently asked questions about mcp-llm-eval
- What is mcp-llm-eval?
- Inferred · not functionally tested: Mcp-llm-eval focuses on automating and standardizing LLM evaluation in CI/CD pipelines for developers and ML engineers. It is catalogued under LLM evaluation & benchmarks on PulseGate.
- Who is mcp-llm-eval for?
- Inferred · not functionally tested: mcp-llm-eval is an open-source project built for machine learning engineers.
- Is mcp-llm-eval free?
- Basis unknown · not verified: Yes — mcp-llm-eval is open source under the MIT license and free to use.
- What platforms does mcp-llm-eval run on?
- Basis unknown · not verified: mcp-llm-eval runs on the command line and API. It can also be self-hosted.
- Is mcp-llm-eval still active?
- PulseGate's liveness check found it on 13 Sep 2026. Its GitHub repository shows 57 commits in the last 90 days.
- What projects are similar to mcp-llm-eval?
- Similar projects tracked by PulseGate include lm-mcp, llmesh-mcp, and mcp-toolgauge.lm-mcpllmesh-mcpmcp-toolgauge
- Who develops mcp-llm-eval?
- mcp-llm-eval is developed by berkayildi.
- When did mcp-llm-eval launch?
- mcp-llm-eval first shipped in 2026.
Similar projects
Closest matches by what these projects do