This repository provides an RKLLM quantized build (w8a8_g128) of the DeepSeek-R1-Distill-Qwen-14B model specifically compiled for RK3588-based single-board computers. It enables local LLM inference on edge devices with a simple Flask server deployment option. The model targets embedded AI applications where cloud connectivity is limited.
In the Foundation models & chat space, DeepSeek R1 Distill Qwen 14B W8a8 G128 Rk3588.rkllm takes a focused approach. It focuses on running capable 14B-scale language models efficiently on low-power Rockchip RK3588 hardware. It is built as an open-source project for developers. DeepSeek R1 Distill Qwen 14B W8a8 G128 Rk3588.rkllm is open source under the Open Source license. It runs on the web and API.
jamescallander builds and maintains DeepSeek R1 Distill Qwen 14B W8a8 G128 Rk3588.rkllm, and the product first shipped in 2024. Development happens publicly on GitHub with 1.6k stars and 1 commits in the last 90 days. It operates in a well-populated space: PulseGate tracks 5 similar tools. Key capabilities include Edge AI Inference, Quantized LLM, and RK3588 Optimization.
Latest indexed changes and source events
jamescallander/DeepSeek-R1-Distill-Qwen-14B_w8a8_g128_rk3588.rkllm verified by the PulseGate indexer