PulseGateCategoriesMethodologyCompanyThe global software index— through the gate this hour
Coverage—in the index
Freshness—newest listing
Cadence—last week average · — today
Index9 markets90 categories · 139 niches
PulseGate

The global index of software taking shape now.

Stores show what passed through a store. Launch sites show what launched there. Catalogs show what entered their catalog. Each sees the market through its own gate. PulseGate reads across them.

FollowGitHubX (Twitter)LinkedIn
Platform
IndexIndex by setCategoriesIndustry UpdatesMethodologySupply IndexData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutTeamDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateBEBench'd
Visit↗
Skip to content
  1. Index›
  2. LLM eval & observability›
  3. Bench'd
← Back to the index
BE

Bench'd

benchd.ai·LLM eval & observability

Bench'd is a benchmark authority for AI memory systems. Its site describes it as a neutral benchmark and a scoreboard for this area, with public leaderboards and a published methodology intended to support reproducible, independently run evaluation.

The service runs memory systems through benchmark protocols that measure recall, temporal correctness, failure traces, and how efficiently past experience improves future performance. It says every run is cryptographically signed and publicly verifiable, and it adds spend attestation through ProofMeter, marked patent pending. The site also distinguishes between verified results and listed or self-reported claims, and notes that self-reported systems are labeled as not independently run by Bench'd. The benchmark coverage includes tracks such as Conversational Memory, Knowledge Brain, and Agent Memory, with examples shown for systems like LlamaIndex Memory, gbrain, and Letta.

Bench'd lists 64 systems indexed and 13 independently scored, and it says 46 systems are awaiting adapters. It presents benchmark index pages, latest signed receipts, scoring models, methodology documentation, and a trust system with defined tiers. The scoring model separates deterministic exact-match scoring for verified results from an LLM-judged nuance score for synthesis and open-ended recall. It also says that eight benchmark specs are published and 23 failure codes are documented.

The site is delivered as a web service with pages for the leaderboard, benchmarks, docs, and blog, and it invites users to claim a system, connect an official endpoint, and verify results against the public harness. It also offers newsletter updates for new benchmark results and methodology changes.

MITWebCLICloud-managed
BBench'd preview
Visit benchd.ai↗

Overview

6 features

Bench'd is a LLM eval & observability project. It focuses on providing independent, reproducible benchmarks for evaluating and comparing AI memory systems. Bench'd is a B2B product aimed at AI researchers and developers. Bench'd is available on the web and the command line.

Bench'd builds and maintains Bench'd, and it first shipped in 2026. Development happens publicly on GitHub with 43 commits in the last 90 days. Key capabilities include AI memory benchmarking, leaderboard, and cryptographic receipts.

Summary written by a language model from the project’s public pages.

  • ✓AI memory benchmarking
  • ✓Leaderboard
  • ✓Cryptographic receipts
  • ✓Open methodology
  • ✓Spend attestation
  • ✓Reproducible results
Tags
ai-benchmarkingmemory-evaluationleaderboardcryptographic-verificationperformance-metrics
AI capabilities
Structured
Inference: Cloud API

Built with & integrations

Framework
Next.js
Hosting
Vercel
AI providers
multipleopenai
Runs on
BrowserCLICloud-managed
Detected from
Next.js
x-nextjs-prerender header · /_next/static/ in the HTML · __next_f in the HTML
openai
bgpt- in the HTML
Vercel
x-vercel-id header · x-vercel-cache header
multiple
bLangChain in the HTML · bLlamaIndex in the HTML

Trust & compliance

License
MIT
Verified signals
✓HTTPS✓GitHub✓Active maintenance

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed26 Jun · 10:46 UTC
    Listing verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about Bench'd

What does Bench'd do?
Bench'd focuses on providing independent, reproducible benchmarks for evaluating and comparing AI memory systems. It is catalogued under LLM eval & observability on PulseGate.
Who should use Bench'd?
Bench'd is a B2B product built for AI researchers and developers.
What platforms does Bench'd run on?
Bench'd runs on the web and the command line.
Is Bench'd still maintained?
The GitHub repository shows 43 commits in the last 90 days.
What projects are similar to Bench'd?
Similar projects tracked by PulseGate include benchgecko, MineBench, and Terminal-Bench.benchgeckoMineBenchTerminal-Bench
Who develops Bench'd?
Bench'd is developed by Bench'd.
How long has Bench'd been around?
Bench'd first shipped in 2026.
Is Bench'd open source?
Bench'd has a public GitHub repository.

At a glance

Pricing
Paid · pricing page detected
Platforms
Cli · Web
Languages
English
Open source
Yes (GitHub)
License
MIT
Built for
AI researchers and developers
Model
B2B
Solves
Providing independent, reproducible benchmarks for evaluating and comparing AI memory systems.

Registered as

GitHub
benchdai/harness
PyPI
benchd-harness

Developer

Solo developer
↗ GitHub

Open source

View on GitHub →
Stars
0
Forks
0
Open issues
0
Last commit
6 Jun 2026
Commits 90d
43
Contributors
1
Authorship
Solo
Default branch
main

Index record

Identity confidence
High · 92.6
Indexed
26 Jun 2026
Lifecycle
Alive
Last seen
26 Jun 2026
Identity audit (12)
Slug
benchd-harness-benchd-ai
Lifecycle last checked
8 Aug 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
26 Jun 2026
Timeline basis
Indexed-at chronology. This listing's first-seen date was written by the catalog backfill, not observed here, so it is not treated as a sighting.
Name from
Derived from the project's own page and URL.
Category from
Assigned by a language model.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model, checked against the page's own declaration.
Canonical URL
https://benchd.ai/

Ship this? Send a correction — no account, and you get a link to follow it.

Similar projects

Closest matches by what these projects do

  • BEbenchgeckobenchgecko.ai
  • MIMineBenchminebench.ai
  • TETerminal-Benchtbench.ai
  • YOYourBenchhuggingface.co
  • INInferenceBenchpypi.org