‹ BackNewsMoE

MoE

Alibaba open-sources Qwen3.8-Flash-Next as an early look at the Qwen 4 architecture
Nvidia says Groq 3 LPX is in full production, hits 3,400 tokens per second on Gemma 4 31B
Stanford professor Percy Liang opens Marin 535B model training to the public
Z.ai founder Jie Tang says bigger parameter counts no longer tell the full story of model strength
HBF standard goes public, but near-term focus stays on SanDisk earnings and NAND pricing
Alibaba
2026-08-04 01:51:01

Alibaba unveils Qwen3.8-Max with self-reported benchmark lead and aggressive token pricing

Alibaba’s Qwen team has introduced Qwen3.8-Max, a new flagship model that the company says scored 86.1 on OSWorld-Verified, ahead of GPT-5.6 Sol Max at 83.2, Fable 5 at 85.0, and Gemini 3.1 Pro at 76.2. The release also included a broader slate of benchmark claims, such as 93.0 on PaperBench and 86.6 on TerminalBench 2.1, alongside positioning the model for long-running autonomous work rather than standard chatbot use. The model uses a mixture-of-experts architecture with 2.4 trillion total parameters and about 95 billion active during inference, built on the Qwen3.5 architecture with a 1 million-token context window. Alibaba also said Qwen3.8-Max is suited for extended coding tasks, desktop software operation, experiment reproduction, and industrial workflows that feed visual input back into a planning loop. Pricing appears to be a central part of the launch. According to QwenCloud pricing cited in the report, Qwen3.8-Max costs $2 per million input tokens and $6 per million output tokens overseas, bringing the combined total to $8 per million tokens. That is below one-third of Claude Opus 5’s combined $30 and below one-quarter of GPT-5.6 Sol standard mode at $35. Still, the benchmarks and capability demonstrations were all disclosed by Alibaba and have not been independently verified, while the company has yet to publish the licensing terms for the promised open-weight release next week on Hugging Face and ModelScope.

1940
Alibaba unveils Qwen3.8-Max with self-reported benchmark lead and aggressive token pricing
Kimi.ai open-sources MoonEP, AgentENV and FlashKDA alongside Kimi K3 release
AMD unveils fully open-source Instella-MoE, a 16B-parameter model aimed at top-tier open models