This is a GGUF-formatted, NVFP4-quantized release of a 27-billion-parameter model (Qwopus 3.6 v2) that includes Multi-Token Prediction (MTP) capabilities. It is designed for efficient local execution via llama.cpp or compatible GGUF runtimes. The model targets users who need high-performance local LLMs on consumer hardware.
Qwopus3.6 27B V2 MTP sits in PulseGate's Foundation models & chat category. It focuses on providing a highly optimized, quantized large language model for local CPU/GPU inference using llama.cpp. Qwopus3.6 27B V2 MTP is an open-source project aimed at AI developers. The project is open source (MIT). The product ships for the command line.
michaelw9999 builds and maintains Qwopus3.6 27B V2 MTP, and the product first shipped in 2026. The project is developed in the open on GitHub with 38 stars and 84 commits in the last 90 days. PulseGate's similarity index places it among 5 comparable tools. Among its 3 catalogued features are GGUF Format, MTP, and NVFP4 Quantization.
Latest indexed changes and source events
michaelw9999/Qwopus3.6-27B-v2-MTP-NVFP4-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.