Qwen3.8-27B-NVFP4: 27B Multimodal: Breaking Memory Limits with NVFP4
unsloth/Qwen3.8-27B-NVFP4
About the project
This is a version of the Qwen3.8-27B-based model quantized to the NVFP4 format. It significantly reduces file size compared to the original model, enabling execution in environments with memory constraints. Distributed under the Apache 2.0 license, it has minimal restrictions on commercial use.
The model card supports multimodal capabilities that go beyond text generation, handling image and video inputs. It automatically recognizes image and video tokens via Jinja templates and strictly validates system messages and tool call structures. If abnormal input formats are detected, exceptions are raised to ensure stability.
The depth of the reasoning process can be adjusted. The reasoning_effort parameter allows control over the thinking process in three stages: xhigh, medium, and low. Set xhigh for complex logical problems and low for simple tasks to efficiently allocate computational resources.
Search and chat are available directly in local apps such as Unsloth Studio. The model can also be tested in a browser without separate installation via HuggingFace Spaces. It has over 650,000 downloads and is actively discussed in the community.
unsloth/Qwen3.8-27B-NVFP4
The original page has no description.
This introduction was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report errors, attribution issues, or removal requests via Contact.