This is a very small test model used internally by the Hugging Face TRL (Transformer Reinforcement Learning) team. It implements a minimal LlamaForCausalLM architecture and is intended for unit testing and CI pipelines rather than actual language generation.
In the Foundation models & chat space, Tiny LlamaForCausalLM takes a focused approach. It focuses on providing a minimal reproducible model for testing Transformer Reinforcement Learning (TRL) library components. Tiny LlamaForCausalLM is an open-source project aimed at developers. The project is open source (Apache-2.0). The product ships for the web, the command line, and API.
It is developed by TRL Internal Testing, and the product first shipped in 2020. The project is developed in the open on GitHub with 18.9k stars and 496 commits in the last 90 days. PulseGate's similarity index places it among 6 comparable tools.
Latest indexed changes and source events
trl-internal-testing/tiny-LlamaForCausalLM-3.2 verified by the PulseGate indexer