Hermes-4-70B-MLX-5bit is a 5-bit quantized version of the Hermes-4 70B model, optimized for the MLX framework on Apple silicon. It includes a custom chat template supporting function calling and extended reasoning modes. The model is designed for local inference with significantly reduced memory requirements while maintaining strong performance.
In the Foundation models & chat space, Hermes 4 70B takes a focused approach. It focuses on running large language models efficiently on Apple silicon hardware with reduced memory usage. Hermes 4 70B is an open-source project aimed at developers. The project is open source (MIT). Hermes 4 70B is available on the web, the command line, and API.
Behind Hermes 4 70B is lmstudio-community, and the product first shipped in 2023. The project is developed in the open on GitHub with 6.4k stars and 21 commits in the last 90 days. Among its 4 catalogued features are Quantized Weights, MLX Optimized, and Function Calling.
Latest indexed changes and source events
lmstudio-community/Hermes-4-70B-MLX-5bit verified by the PulseGate indexer
Other apps tracked under the same category.