🔥 Qwen3.8-Flash-Next beats models 10x its size
Alibaba just previewed its Qwen4 architecture, and the numbers are brutal for competitors.
The model activates only 6 of 125B parameters per token, yet tops DeepSeek-V4-Flash and Claude Opus 4.6 on coding benchmarks.
Training cost: one-ninth of rivals.
This is the efficiency war going nuclear, not the capability war.
Watch compute prices crater as Chinese labs make "bigger is better" look like a scam.

August 26, 2026 148 1