AI Briefing
KO

Qwen3.8-27B-NVFP4: 27B Multimodal: Breaking Memory Limits with NVFP4

unsloth/Qwen3.8-27B-NVFP4

·2026.08.20 09:00

This is a version of the Qwen3.8-27B-based model quantized to the NVFP4 format. It significantly reduces file size compared to the original model, enabling execution in environments with memory constraints. Distributed under the Apache 2.0 license, it has minimal restrictions on commercial use.

The model card supports multimodal capabilities that go beyond text generation, handling image and video inputs. It automatically recognizes image and video tokens via Jinja templates and strictly validates system messages and tool call structures. If abnormal input formats are detected, exceptions are raised to ensure stability.

The depth of the reasoning process can be adjusted. The reasoning_effort parameter allows control over the thinking process in three stages: xhigh, medium, and low. Set xhigh for complex logical problems and low for simple tasks to efficiently allocate computational resources.

Search and chat are available directly in local apps such as Unsloth Studio. The model can also be tested in a browser without separate installation via HuggingFace Spaces. It has over 650,000 downloads and is actively discussed in the community.

HuggingFace
HuggingFace model

unsloth/Qwen3.8-27B-NVFP4

The original page has no description.

This introduction was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report errors, attribution issues, or removal requests via Contact.