internlm2-1_8b-reward is a 1.8 billion parameter reward model from the InternLM2 family. It is designed to evaluate and score text outputs for quality, helpfulness, or alignment. It is primarily used by AI researchers and developers building or fine-tuning language models with preference optimization techniques.
In the Other AI space, Internlm2 1 8b Reward takes a focused approach. It focuses on scoring and ranking model outputs to support reinforcement learning from human feedback (RLHF) pipelines. Internlm2 1 8b Reward is an open-source project aimed at developers. The project is open source (Apache-2.0). It runs on the web and API.
InternLM builds and maintains Internlm2 1 8b Reward, and the product first shipped in 2023. The project is developed in the open on GitHub with 7.2k stars.
Latest indexed changes and source events
internlm/internlm2-1_8b-reward verified by the PulseGate indexer
Other apps tracked under the same category.