This repository hosts a 6-bit quantized version of the Hermes-4-70B model in MLX format, optimized for Apple silicon (M-series chips). It includes advanced system prompts for deep thinking and tool use, enabling high-quality local inference on Macs without requiring high-end GPUs.
In the Foundation models & chat space, Hermes 4 70B takes a focused approach. It focuses on running a high-parameter reasoning model efficiently on Apple Mac hardware with MLX. Hermes 4 70B is an open-source project aimed at developers. The project is open source (MIT). The product ships for the web, the command line, and API.
lmstudio-community builds and maintains Hermes 4 70B, and the product first shipped in 2023. The project is developed in the open on GitHub with 6.4k stars and 21 commits in the last 90 days. Among its 3 catalogued features are quantized, MLX Format, and reasoning.
Latest indexed changes and source events
lmstudio-community/Hermes-4-70B-MLX-6bit verified by the PulseGate indexer
Other apps tracked under the same category.