TinyLlama-1.1B-Chat-v1.0-GGUF is a quantized GGUF distribution of the TinyLlama 1.1B chat model, created by TheBloke for efficient local execution. It enables users to run a small but capable language model on CPUs or low-end GPUs using tools like llama.cpp. This format is popular among developers and hobbyists who want to experiment with LLMs without high-end hardware.
TinyLlama 1.1B Chat sits in PulseGate's Quantised & converted weights category. It focuses on running a capable small language model on consumer hardware with minimal resource requirements through quantization. TinyLlama 1.1B Chat is an open-source project aimed at developers and enthusiasts. The project is open source (MIT). It ships for the command line, and it can be self-hosted.
It is developed by TheBloke, and it first shipped in 2023. The project is developed in the open on GitHub with 121k stars and 1.2k commits in the last 90 days. Among its 4 catalogued features are GGUF Quantization, Local Chat Interface, and Lightweight LLM.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do