AI Briefing
KO

Qwen3.8-Flash-Next-GSQ-RCO-GGUF: Mixed-Precision GGUF of Qwen3.8 Flash Next with GSQ and RCO

ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF

·2026.09.19 23:46

This is a GGUF distribution of the Qwen3.8 Flash Next model, quantized using GSQ and RCO techniques. It maintains the Mixture of Experts architecture while providing checkpoints at various bit widths, including IQ2_XS, IQ3_XXS, and Q2_0.

It supports multimodal capabilities for processing both images and text. A BF16 mmproj file is included, enabling integration between the vision encoder and the language model.

It is ready for use in local inference environments such as llama.cpp and vLLM. Released under the Apache 2.0 license, it can be applied to experimental and commercial environments without restrictions.

HuggingFace
HuggingFace model

ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF

The original page has no description.

image-text-to-text

This introduction was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report errors, attribution issues, or removal requests via Contact.