Qwen3.8-Flash-Next-GSQ-RCO-GGUF: Mixed-Precision GGUF of Qwen3.8 Flash Next with GSQ and RCO
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF
About the project
This is a GGUF distribution of the Qwen3.8 Flash Next model, quantized using GSQ and RCO techniques. It maintains the Mixture of Experts architecture while providing checkpoints at various bit widths, including IQ2_XS, IQ3_XXS, and Q2_0.
It supports multimodal capabilities for processing both images and text. A BF16 mmproj file is included, enabling integration between the vision encoder and the language model.
It is ready for use in local inference environments such as llama.cpp and vLLM. Released under the Apache 2.0 license, it can be applied to experimental and commercial environments without restrictions.
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF
The original page has no description.
image-text-to-text
This introduction was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report errors, attribution issues, or removal requests via Contact.





