This is a GPTQ 4-bit quantized version of TinyLlama-1.1B-Chat-v1.0 by TheBloke. It provides an efficient way to run the 1.1 billion parameter chat model locally with significantly reduced memory requirements while maintaining good performance.
TinyLlama 1.1B Chat sits in PulseGate's Foundation models & chat category. It focuses on running a capable small language model for chat on resource-constrained devices through 4-bit GPTQ quantization. TinyLlama 1.1B Chat is an open-source project aimed at Local LLM users and developers. The project is open source (AGPL-3.0). The product ships for the web, the command line, and API.
It is developed by TheBloke, and the product first shipped in 2022. The project is developed in the open on GitHub with 47.5k stars and 141 commits in the last 90 days. PulseGate's similarity index places it among 5 comparable tools. Among its 3 catalogued features are GPTQ quantization, chat format support, and small footprint.
Latest indexed changes and source events
TheBloke/TinyLlama-1.1B-Chat-v1.0-GPTQ verified by the PulseGate indexer
Other apps tracked under the same category.