Skip to content
Back to the index

kvfit

PyPIInfrastructure

PulseGate's liveness check found it on 14 Sep 2026; it is registered on GitHub and PyPI and has been in the index since 21 Jul 2026. How this is checked

kvfit is a command-line utility that helps developers determine whether a Hugging Face LLM will fit on their GPU or DGX hardware. It accounts for model weights, KV cache or recurrent state, context length, tensor and data parallelism, and provides measured calibration. Useful for capacity planning and inference optimization in LLM deployment workflows.

Inferred · not functionally tested

Open SourceMITCLI
Visit PyPI

Overview

4 features

Purpose: Determining if a large language model will fit in available GPU or DGX memory before attempting to load it.

Inferred · not functionally tested

Audience: developers

Inferred · not functionally tested

Functions: Unknown

Interfaces: API: unknown · MCP: unknown · CLI: indicated (inferred, not tested) · Self-hosting: unknown

Recorded constraints: pricing: open_source · license: MIT · platforms: CLI · deployment: cli

Constraint provenance is unknown; confirm requirements with the publisher.

Record sources: pypi.org · github.com. These links do not verify the individual claims.

kvfit is an Inference & model serving project. Inferred · not functionally tested: It focuses on determining if a large language model will fit in available GPU or DGX memory before attempting to load it. Inferred · not functionally tested: It is built as an open-source project for developers. Basis unknown · not verified: kvfit is open source under the MIT license. Basis unknown · not verified: kvfit is available on the command line.

Teilomillet builds and maintains kvfit, and it first shipped in 2026. Development happens publicly on GitHub with 4 commits in the last 90 days. Inferred · not functionally tested: Key capabilities include GPU Memory Calculator, KV Cache Planning, and Model Fit Analysis.

Summary written by a language model from the project’s public pages.

Tasks: Inferred · not functionally tested

  • GPU Memory Calculator
  • KV Cache Planning
  • Model Fit Analysis
  • Hugging Face Integration

Topics: Inferred · not functionally tested

Tags
gpu-memory-calculatorllm-inferencemodel-servingkv-cache

JSON profile · Text profile · Access guide

Built with & integrations

Runs on
CLI

Trust & compliance

License
MIT
Public signals
HTTPSOpen SourceGitHubActive maintenance

Indexing history

2

What PulseGate has recorded for this listing

  1. Indexed22 Sep · 18:17 UTC
    kvfit seen via PyPI Fresh Feed
    Source: PyPI Fresh Feed · Open
  2. Indexed21 Jul · 18:41 UTC
    kvfit seen via PyPI Fresh Feed
    Source: PyPI Fresh Feed · Open

Frequently asked questions about kvfit

What does kvfit do?
Inferred · not functionally tested: Kvfit focuses on determining if a large language model will fit in available GPU or DGX memory before attempting to load it. It is catalogued under Inference & model serving on PulseGate.
Who should use kvfit?
Inferred · not functionally tested: kvfit is an open-source project built for developers.
Is kvfit free?
Basis unknown · not verified: Yes — kvfit is open source under the MIT license and free to use.
What platforms does kvfit run on?
Basis unknown · not verified: kvfit runs on the command line.
Is kvfit still maintained?
PulseGate's liveness check found it on 14 Sep 2026. Its GitHub repository shows 4 commits in the last 90 days.
What projects are similar to kvfit?
Similar projects tracked by PulseGate include ModelFit, LLM Model VRAM Calculator, and quantfit.ModelFitLLM Model VRAM Calculatorquantfit
Who develops kvfit?
kvfit is developed by Teilomillet.
How long has kvfit been around?
kvfit first shipped in 2026.

Similar projects

Closest matches by what these projects do