Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency"
The Decoderen

Alibaba's Qwen team is previewing the Qwen4 architecture with Qwen3.8-Flash-Next, a mixture-of-experts model that activates just 6 out of 125 billion parameters per token. At one-ninth the training cost, it beats much larger competitors like DeepSeek-V4-Flash and Claude Opus 4.6 on coding and office benchmarks, adding more pricing pressure on OpenAI and Anthropic. The article Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency" appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- OpenAI
- Anthropic
- DeepSeek
- Verktyg
Related AI news
- Preparing data for supervised fine-tuning Part 1: Formatting and qualityAWS Machine Learning · August 26, 2026
- Interviews with OpenAI leaders, employees, and others on the company's reboot; Sam Altman says OpenAI would have a system he would call AGI by the end of 2026 (Alex Heath/Time)Techmeme · August 26, 2026
- Bengaluru-based Runable, whose AI agents let small businesses find customers, run ad campaigns, and more, raised a $21M Series A at a $65M post-money valuation (Jagmeet Singh/TechCrunch)Techmeme · August 26, 2026
- Orchestration is the new challenge for CX in the age of AI agentsVentureBeat AI · August 26, 2026
- Apples Foldable: Leaks zeigen Mainboards des m�glichen iPhone UltraGolem.de · August 26, 2026
- Nuoret saavat kriisitukea Deepfake-tapauksen vuoksi LappeenrannassaYle Uutiset · August 26, 2026