EAGLE-LLaMA3-Instruct-8B is an instruction-tuned 8 billion parameter model based on Llama 3 that incorporates EAGLE speculative decoding techniques for faster inference. It supports popular libraries including Transformers and vLLM and is available on Hugging Face. The model is built for developers seeking high-speed text generation while maintaining the quality of Llama 3 instruct capabilities.
EAGLE LLaMA3 Instruct 8B sits in PulseGate's Foundation models & chat category. High latency in autoregressive text generation for instruction-tuned LLMs. It is built as an open-source project for developers. EAGLE LLaMA3 Instruct 8B is open source under the Open Source license. EAGLE LLaMA3 Instruct 8B is available on the web, the command line, and API.
Behind EAGLE LLaMA3 Instruct 8B is yuhuili, and the product first shipped in 2023. Development happens publicly on GitHub with 2.5k stars. Key capabilities include text generation, vLLM support, and transformers integration.
Latest indexed changes and source events
yuhuili/EAGLE-LLaMA3-Instruct-8B verified by the PulseGate indexer
Other apps tracked under the same category.