According to Artificial Analysis, Alibaba's newly released Qwen3.8 Max scores 56 on the Artificial Analysis Intelligence Index, a 10-point improvement over its predecessor Qwen3.7 Max's 46, tying Claude Opus 4.8 (max, 56), and trailing only Kimi K3 (max, 57) among Chinese labs, ahead of GLM-5.2 (max, 51).
On GDPval-AA, Qwen3.8 Max achieves 1739 Elo, a 468-point improvement over the previous generation, surpassing Kimi K3 (1685) and roughly matching Claude Fable 5 (1743) and GPT-5.6 Sol (max, 1730), trailing only Claude Opus 5 (max, 1852).
Compared to its predecessor, Qwen3.8 Max shows improvements in agentic evaluations, scientific reasoning, and coding: Terminal-Bench v2.1 up 6 points, CritPt up 7, SciCode up 4, HLE up 3, while GPQA remains unchanged.