See exactly where Qwen 3.8 outscores a frontier closed model, and what it really takes to run it.
What's inside
by Divjot Sahni · EZYE Consulting · ezye.com.au
The Headline Result
Qwen 3.8 runs with only 6B active parameters
It beats Claude Opus 4.6 on two real coding benchmarks
The win is not marginal, it's a clean beat on both
But the fine print changes who can actually use it...
The Numbers Side by Side
SWE-bench Pro: Qwen 62.5 vs Claude Opus 4.6 at 53.4
Office-agent tasks: Qwen 73.9 vs Claude Opus 4.6 at 68.2
Both are real-world agent benchmarks, not toy tests
The catch that stops most people from running it: ...
Free access — enter your email below