llm-speed
PulseGate's liveness check found it on 13 Sep 2026; it is registered on GitHub and PyPI and has been in the index since 27 Jun 2026. How this is checked
llm-speed is an open-source CLI tool that benchmarks the inference speed of large language models (LLMs) across various hardware setups and hosted APIs. It provides reproducible, community-verified results for AI researchers and developers seeking to optimize model performance.
Inferred · not functionally tested
Overview
5 featuresPurpose: Measuring and comparing LLM inference speed across different hardware and API backends.
Inferred · not functionally tested
Audience: AI researchers and developers
Inferred · not functionally tested
Functions: analytics, monitoring
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: unknown
Recorded constraints: pricing: open_source · license: Apache-2.0 · platforms: CLI · deployment: cli
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: llm-speed.com · github.com. These links do not verify the individual claims.
In the CLI tools & terminal space, llm-speed takes a focused approach. Inferred · not functionally tested: It focuses on measuring and comparing LLM inference speed across different hardware and API backends. Inferred · not functionally tested: llm-speed is an open-source project aimed at AI researchers and developers. Basis unknown · not verified: llm-speed is open source under the Apache-2.0 license. Basis unknown · not verified: It ships for the command line.
llm-speed first shipped in 2026. The project is developed in the open on GitHub with 4 commits in the last 90 days. Inferred · not functionally tested: Among its 5 catalogued features are LLM benchmarking, token speed measurement, and hardware detection.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- LLM benchmarking
- Token speed measurement
- Hardware detection
- Open source
- Reproducible methodology
Topics: Inferred · not functionally tested
Built with & integrations
- Next.js
- x-nextjs-prerender header · /_next/static/ in the HTML · __next_f in the HTML
- local_oss
- bllama in the HTML · bvllm in the HTML
- Cloudflare
- cf-ray header · cf-cache-status header
- meta_llama
- llama- in the HTML
Trust & compliance
Indexing history
What PulseGate has recorded for this listing
Frequently asked questions about llm-speed
- What does llm-speed do?
- Inferred · not functionally tested: Llm-speed focuses on measuring and comparing LLM inference speed across different hardware and API backends. It is catalogued under CLI tools & terminal on PulseGate.
- Who should use llm-speed?
- Inferred · not functionally tested: llm-speed is an open-source project built for AI researchers and developers.
- Is llm-speed free?
- Basis unknown · not verified: Yes — llm-speed is open source under the Apache-2.0 license and free to use.
- What platforms does llm-speed run on?
- Basis unknown · not verified: llm-speed runs on the command line.
- Is llm-speed still maintained?
- PulseGate's liveness check found it on 13 Sep 2026. Its GitHub repository shows 4 commits in the last 90 days.
- What projects are similar to llm-speed?
- Similar projects tracked by PulseGate include llm-benchmark-runner, llm-cal, and llm-inspector.llm-benchmark-runnerllm-calllm-inspector
- How long has llm-speed been around?
- llm-speed first shipped in 2026.
- Is llm-speed open source?
- Basis unknown · not verified: Yes — llm-speed is open source under the Apache-2.0 license, developed on GitHub.
Similar projects
Closest matches by what these projects do