This repository provides GGUF quantized versions of the umt5-xxl encoder model for efficient local inference. Multiple quantization levels (Q2 through Q8, F16, F32) are available. The model is suitable for multilingual text encoding tasks and can be used with llama.cpp and other GGUF-compatible runtimes.
In the Foundation models & chat space, Umt5 Xxl Encoder takes a focused approach. It focuses on running large multilingual T5 encoder models efficiently on consumer hardware using quantized GGUF format. It is built as an open-source project for developers. Umt5 Xxl Encoder is open source under the MIT license. Umt5 Xxl Encoder is available on the web, the command line, and API.
Behind Umt5 Xxl Encoder is city96, and the product first shipped in 2023. Development happens publicly on GitHub with 121k stars and 1.2k commits in the last 90 days. Key capabilities include GGUF Quantization, Encoder Model, and multilingual.
Latest indexed changes and source events
city96/umt5-xxl-encoder-gguf verified by the PulseGate indexer
Other apps tracked under the same category.