Alibaba released Qwen3.8-Max on Monday, describing it as the most capable model in its Qwen family to date, with 2.4 trillion total parameters and 95 billion active parameters.
The model is
available through QwenCloud, with open weights planned for release next week — marking the first time Alibaba has open-sourced a model at this scale.
Each query can draw on a context window of up to 1 million tokens — enough to cover approximately 750,000 words of input. Alibaba described its capabilities as spanning coding, real-world work tasks, long-horizon autonomous operation, and multimodal understanding of documents, video,
and images.
Alibaba published benchmark comparisons showing Qwen3.8-Max performing at comparable levels to Anthropic's Fable 5 on several coding and general agent tasks, and scoring above it on some multimodal and document benchmarks. On PaperBench, Qwen3.8-Max posted a score of 93.0 against Fable 5's 88.8. On the general capability benchmark IFBench, it scored 82.8 against Fable 5's 63.5. Fable 5 led on several other measures, including SWE-bench Pro and
most visual agent tasks. |