Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency"

The Decoder The Decoder

https://the-decoder.com/wp-content/uploads/2026/07/qwen_logo-1.png" style="height: auto; margin-bottom: 10px;" width="2048" />


Alibaba's Qwen team is previewing the Qwen4 architecture with Qwen3.8-Flash-Next, a mixture-of-experts model that activates just 6 out of 125 billion parameters per token.

At one-ninth the training cost, it beats much larger competitors like DeepSeek-V4-Flash and Claude Opus 4.6 on coding and office benchmarks, adding more pricing pressure on OpenAI and Anthropic.


The article https://the-decoder.com/alibaba-releases-qwen3-8-flash-next-targeting-ultimate-cost-efficiency/">Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency" appeared first on https://the-decoder.com">The Decoder.

Read full article at The Decoder →