PulseGateIndexCategoriesUpdatesLive intelligence on AI-era software— indexed in the last hour
Coverage—projects tracked
Freshness—newest entry
Cadence—last week average · — today
Index9 markets86 categories
PulseGate

Live intelligence on the AI-era software market — apps, models, agents and infrastructure.

Most of it never reaches an official store. PulseGate maps the whole market — every category, worldwide — and measures how it moves: what’s launching, what’s gaining, what’s going quiet.

FollowGitHubX (Twitter)LinkedIn
Platform
All AppsFull IndexCategoriesIndustry UpdatesData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateKMKVzap Mlp Llama 3.1 8B Instruct
Visit↗
Skip to content
  1. Index›
  2. Other AI›
  3. KVzap Mlp Llama 3.1 8B Instruct
← Back to the index
KM

KVzap Mlp Llama 3.1 8B Instruct

huggingface.co·Infrastructure·🇺🇸

nvidia/KVzap-mlp-Llama-3.1-8B-Instruct is a variant of the Llama 3.1 8B model that incorporates KVzap, a method for pruning the key-value cache using a lightweight MLP to predict importance scores. This accelerates both prefilling and decoding phases of LLM inference while maintaining model quality. It is hosted on Hugging Face and can be used with the Transformers library.

Open SourceApache-2.0WebAPI
KKVzap Mlp Llama 3.1 8B Instruct preview
Visit huggingface.co↗
1.2kstars
167forks
4features
2024since

Overview

4 features

KVzap Mlp Llama 3.1 8B Instruct is an Other AI project. It focuses on reducing the computational cost and memory usage of large language model inference through intelligent KV cache pruning. KVzap Mlp Llama 3.1 8B Instruct is an open-source project aimed at machine learning engineers and researchers. KVzap Mlp Llama 3.1 8B Instruct is open source under the Apache-2.0 license. KVzap Mlp Llama 3.1 8B Instruct is available on the web and API.

NVIDIA builds and maintains KVzap Mlp Llama 3.1 8B Instruct, and it first shipped in 2024. The project is developed in the open on GitHub with 1.2k stars and 14 commits in the last 90 days. Key capabilities include KV Cache Pruning, Fast Inference, and llama-based. It exposes integrations via a public API.

Summary written by a language model from the project’s public pages.

  • ✓KV Cache Pruning
  • ✓Fast Inference
  • ✓Llama-based
  • ✓Instruction Tuned
Tags
llama-3kv-cache-pruninginference-optimizationnvidia-model
AI capabilities
Inference: LocalWeights: Open

Built with & integrations

Hosting
AWS
AI providers
meta_llama
Connectors
API
Runs on
BrowserAPI-only
Detected from
AWS
x-amz-cf-id header · x-amz-cf-pop header · via header
meta_llama
llama- in the HTML

Trust & compliance

License
Apache-2.0
Verified signals
✓HTTPS✓Privacy Policy✓Terms of Service✓Open Source✓Active maintenance
Legal
Privacy Policy →Terms of Service →

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed5 Aug · 09:38 UTC
    nvidia/KVzap-mlp-Llama-3.1-8B-Instruct verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about KVzap Mlp Llama 3.1 8B Instruct

What does KVzap Mlp Llama 3.1 8B Instruct do?
KVzap Mlp Llama 3.1 8B Instruct focuses on reducing the computational cost and memory usage of large language model inference through intelligent KV cache pruning. It is catalogued under Other AI on PulseGate.
Who should use KVzap Mlp Llama 3.1 8B Instruct?
KVzap Mlp Llama 3.1 8B Instruct is an open-source project built for machine learning engineers and researchers.
Is KVzap Mlp Llama 3.1 8B Instruct free?
Yes — KVzap Mlp Llama 3.1 8B Instruct is open source under the Apache-2.0 license and free to use.
What platforms does KVzap Mlp Llama 3.1 8B Instruct run on?
KVzap Mlp Llama 3.1 8B Instruct runs on the web and API.
Is KVzap Mlp Llama 3.1 8B Instruct still active?
The GitHub repository shows 14 commits in the last 90 days.
What projects are similar to KVzap Mlp Llama 3.1 8B Instruct?
Similar projects tracked by PulseGate include Llama 3.1 8B Instruct, Llama 3.1 8B Instruct FP8 KV, and Llama 3.3 70B Instruct.Llama 3.1 8B InstructLlama 3.1 8B Instruct FP8 KVLlama 3.3 70B Instruct
Who makes KVzap Mlp Llama 3.1 8B Instruct?
KVzap Mlp Llama 3.1 8B Instruct is developed by NVIDIA, based in the United States.
When did KVzap Mlp Llama 3.1 8B Instruct launch?
KVzap Mlp Llama 3.1 8B Instruct first shipped in 2024.

At a glance

Pricing
Open Source
Platforms
Web
Languages
English
License
Apache-2.0
Built for
machine learning engineers and researchers
Model
Open source
Solves
Reducing the computational cost and memory usage of large language model inference through intelligent KV cache pruning.

Registered as

Hugging Face
nvidia/KVzap-mlp-Llama-3.1-8B-Instruct

Developer

🇺🇸NVIDIA
Community-driven

Open source

Stars
1,155
Forks
167
Open issues
5
Last commit
31 Jul 2026
Commits 90d
14
Contributors
29
Authorship
Community-driven
Default branch
main
Latest release
v0.5.4 · 2 Jul 2026

Live coverage

Identity confidence
Medium · 70.4
Indexed
5 Aug 2026
Lifecycle
Alive
Last seen
5 Aug 2026
Identity audit (13)
Slug
nvidia-kvzap-mlp-llama-3-1-8b-instruct-huggingface-co
Lifecycle state recorded
5 Aug 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
5 Aug 2026
Timeline basis
Indexed-at chronology. This listing's first-seen date was written by the catalog backfill, not observed here, so it is not treated as a sighting.
Name from
Derived from the project's own page and URL.
Category from
Derived from the model's own Hugging Face metadata.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model from page content.
Last updated
5 Aug 2026
Canonical URL
https://huggingface.co/nvidia/KVzap-mlp-Llama-3.1-8B-Instruct

Ship this? Send a correction — no account, and you get a link to follow it.

Similar apps

Closest matches by what these projects do

  • L3Llama 3.1 8B Instructhuggingface.co
  • L3Llama 3.1 8B Instruct FP8 KVhuggingface.co
  • L3Llama 3.3 70B Instructhuggingface.co
  • L3Llama 3.1 8B Instructhuggingface.co
  • L3Llama 3 3 Nemotron Super 49B V1 5huggingface.co
  • L3Llama 3.3 70B Instructhuggingface.co