Swift-Qwen3.8-27B-GGUF: IQ2~Q8 Quantization Set for 27B-class Vision Model
ukisai/Swift-Qwen3.8-27B-GGUF
About the project
This model is converted to GGUF format while maintaining the inference performance of Swift-Qwen3.8-27B. It supports multimodal capabilities, accepting both text and image inputs to generate responses, and can be directly deployed in local execution environments such as llama.cpp or Ollama.

A total of 28 quantization versions are provided, ranging from F16 to IQ2_XXS. Depending on hardware specifications, users can choose between the memory-efficient IQ series or the quality-focused Q series. The visual encoder (mmproj) files are also included.
The imatrix quantization method is applied to minimize performance degradation even at low file sizes. It supports thinking mode and tool calling, with adjustable reasoning depth via the reasoning_effort setting. The model follows the swift-open-license-1.0 license.
ukisai/Swift-Qwen3.8-27B-GGUF
The original page has no description.
image-text-to-text
This introduction was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report errors, attribution issues, or removal requests via Contact.