What was claimed
Qwen3.8-Flash 125B MoE model can now be run locally and outperforms Claude-Opus-4.6 (Max). Run on 75GB RAM via Unsloth GGUFs.
Our verdict
Needs cautionQwen3.8-Flash-Next outperforms Claude Opus 4.6 Max on most practical coding and agent evaluations, but it is not a clean sweep—Opus 4.6 Max still leads on HLE (40.0 vs. 35.9). The claim of outperformance is benchmark-specific and vendor-reported, not universal. The actual model name is Qwen3.8-Flash-Next, not simply 'Qwen3.8-Flash'. While a production version Qwen3.8-Flash is available via QwenCloud API, the open-weight local model is specifically the 'Next' variant. The claim conflates the two. (Only 1 of 3 AI systems responded.)
Check your own claim
Paste any statement, headline, or AI answer — 3 independent AIs verify it in seconds, with sources.
Key findings
Model is called 'Qwen3.8-Flash'
Outperforms Claude-Opus-4.6 (Max)
Run on 75GB RAM via Unsloth GGUFs
Qwen3.8-Flash 125B MoE model can now be run locally