Qwen3.8-Flash: 6B-Active Next-Gen MoE, a Qwen4 Preview
125B transformer + 51B N-gram embeddings, only 6B active. SWE-bench Pro beats Claude Opus 4.6 by 9.1 points, JobBench by ~20; native multimodal (AndroidWorld +22.5, MathVision +25.1, ERQA +31.5 vs Opus 4.6). QSA sparse attention + Gated Residual cut long-context cost ~8x on cache hits.