AI Briefing
KO

RTX 6000 Pro, Qwen3.6 27B Benchmark

·2026.05.02 08:15

Key point

The RTX 6000 Pro was significantly faster than the 4080 Super at Qwen3.6 27B inference.

Details

Running Qwen3.6 27B on an RTX 6000 Pro rig showed a huge performance gap compared to the 4080 Super.

  • 4080 Super: Qwen3.6 27B Q2 quantization, about 6 tok/s, TTFT about 60 seconds
  • RTX 6000 Pro: Qwen3.6 27B Q8 XL, 67 tok/s, TTFT about 1 second

Based on initial LM Studio benchmarks, token generation was about 10x faster and prompt processing speed was about 60x faster. No additional tuning has been done yet.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.