FireTofu
Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency"

Technology · en

Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency"

The Decoder · Aug 26, 2026, 2:40 PM UTC

Alibaba's Qwen team is previewing the Qwen4 architecture with Qwen3.8-Flash-Next, a mixture-of-experts model that activates just 6 out of 125 billion parameters per token. At one-ninth the training cost, it beats much larger competitors like DeepSeek-V4-Flash and Claude Opus 4.…