Sugoi-14B-Ultra is a 14 billion parameter instruction-tuned language model converted to GGUF format for efficient CPU and GPU inference using tools like llama.cpp. It includes support for tool calling and follows a Qwen-style chat template. The model is designed for high-quality local deployment across various hardware configurations.
In the Foundation models & chat space, Sugoi 14B Ultra takes a focused approach. It focuses on running high-performance 14B parameter models locally with flexible quantization options. Sugoi 14B Ultra is an open-source project aimed at local LLM users and application developers. The project is open source (Open Source). It runs on the web, the command line, and API.
sugoitoolkit builds and maintains Sugoi 14B Ultra, and the product first shipped in 2024. Among its 3 catalogued features are Large Language Model, GGUF Quantization, and Tool Calling Support.
Latest indexed changes and source events
sugoitoolkit/Sugoi-14B-Ultra-GGUF verified by the PulseGate indexer
Other apps tracked under the same category.