This is a very small causal language model created by the Hugging Face TRL team for internal testing and CI purposes. It is not intended for production use but serves as a lightweight example for developers working with the Transformers Reinforcement Learning library.
Tiny Lfm2ForCausalLM is a Foundation models & chat project. It focuses on providing a minimal reproducible model for testing reinforcement learning and fine-tuning code in the TRL library. Tiny Lfm2ForCausalLM is an open-source project aimed at developers. The project is open source (Apache-2.0). Tiny Lfm2ForCausalLM is available on the web, the command line, and API, and it can be self-hosted.
It is developed by trl-internal-testing, and it first shipped in 2020. Development happens publicly on GitHub with 19k stars and 547 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 17 similar projects. It exposes integrations via a public API.
Summary written by a language model from the project’s public pages.
What PulseGate has recorded for this listing
Closest matches by what these projects do