vapor-ram is an Apache-2.0 local LLM inference server with a command-line interface and dashboard for running Google Gemma models. It is intended for developers and technical users who want to serve models on their own hardware.
vapor-ram sits in PulseGate's Other models category. It focuses on running and serving a local Gemma language model without relying on a hosted inference provider. vapor-ram is an open-source project aimed at developers and technical users. The project is open source (Apache-2.0). It ships for the command line and the web, and it can be self-hosted.
Behind vapor-ram is sudsarkar13, and it first shipped in 2026. Development happens publicly on GitHub with 132 commits in the last 90 days. Among its 4 catalogued features are local inference, CLI interface, and web dashboard.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Same category — not a similarity match