PulseGateIndexCategoriesUpdatesLive intelligence on AI-era software— indexed in the last hour
Coverage—projects tracked
Freshness—newest entry
Cadence—last week average · — today
Index9 markets86 categories
PulseGate

Live intelligence on the AI-era software market — apps, models, agents and infrastructure.

Most of it never reaches an official store. PulseGate maps the whole market — every category, worldwide — and measures how it moves: what’s launching, what’s gaining, what’s going quiet.

FollowGitHubX (Twitter)LinkedIn
Platform
All AppsFull IndexCategoriesIndustry UpdatesData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateMEmendmark-evals
Visit↗
Skip to content
  1. Index›
  2. LLM eval & observability›
  3. mendmark-evals
← Back to the index
ME

mendmark-evals

PyPI·Infrastructure

Mendmark-evals is a Python package for performing mutation testing on evaluation suites used for LLM-based agents. It helps developers assess how well their test frameworks detect failures and measure the quality of autonomous agent implementations.

Open SourceMITCLI
Visit PyPI↗

Overview

3 features

mendmark-evals sits in PulseGate's LLM eval & observability category. It focuses on evaluating the robustness and reliability of AI agent testing frameworks. It is built as an open-source project for AI developers. The project is open source (MIT). It ships for the command line.

It is developed by Daniel Gaskins, and it first shipped in 2026. Development happens publicly on GitHub with 14 commits in the last 90 days. Key capabilities include mutation testing, agent evaluation, and test suite analysis.

Summary written by a language model from the project’s public pages.

  • ✓Mutation testing
  • ✓Agent evaluation
  • ✓Test suite analysis
Tags
mutation-testingagent-evaluationllm-testingeval-framework

Built with & integrations

Runs on
CLI

Trust & compliance

License
MIT
Verified signals
✓HTTPS✓Open Source✓GitHub✓Active maintenance

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed3 Aug · 18:46 UTC
    mendmark-evals verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about mendmark-evals

What does mendmark-evals do?
Mendmark-evals focuses on evaluating the robustness and reliability of AI agent testing frameworks. It is catalogued under LLM eval & observability on PulseGate.
Who is mendmark-evals for?
mendmark-evals is an open-source project built for AI developers.
Does mendmark-evals have a free plan?
Yes — mendmark-evals is open source under the MIT license and free to use.
What platforms does mendmark-evals run on?
mendmark-evals runs on the command line.
Is mendmark-evals still maintained?
The GitHub repository shows 14 commits in the last 90 days.
What are alternatives to mendmark-evals?
Similar projects tracked by PulseGate include ase-python, constraintloop, and OpenCode Data.ase-pythonconstraintloopOpenCode Data
Who develops mendmark-evals?
mendmark-evals is developed by Daniel Gaskins.
When did mendmark-evals launch?
mendmark-evals first shipped in 2026.

At a glance

Platforms
Cli
Languages
English
Open source
Yes (GitHub)
License
MIT
First seen
3 Aug 2026
Built for
AI developers
Model
Open source
Solves
Evaluating the robustness and reliability of AI agent testing frameworks.

Registered as

GitHub
danielgaskins/mendmark
PyPI
mendmark-evals

Developer

Daniel Gaskins
Solo developer
↗ GitHub

Open source

View on GitHub →
Stars
0
Forks
0
Open issues
1
Last commit
3 Aug 2026
Commits 90d
14
Contributors
1
Authorship
Solo
Default branch
main
Latest release
v0.4.0 · 3 Aug 2026

Live coverage

Identity confidence
Low · 64
Indexed
3 Aug 2026
Lifecycle
Alive
First seen
Aug 2026
Last seen
3 Aug 2026
Identity audit (13)
Slug
mendmark-evals-pypi-org
Lifecycle state recorded
3 Aug 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
3 Aug 2026
Timeline basis
Indexed-at chronology (no inferred launch/funding milestones).
Name from
Written by a language model from the project's public pages.
Category from
Assigned by a language model.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model from page content.
Last updated
3 Aug 2026
Canonical URL
https://pypi.org/project/mendmark-evals

Ship this? Send a correction — no account, and you get a link to follow it.

Also in LLM eval & observability

Same category — not a similarity match

  • ASase-pythongithub.com
  • COconstraintlooppypi.org
  • ODOpenCode Dataopencode.ai
  • COCOHESIONcohesionauth.com
  • TOTokenTelemetrytokentelemetry.com
  • DODOOMCLOCKdoomclock.app