Alibaba’s Qwen3.8 Max has landed a score of 56 on the Artificial Analysis Intelligence Index, the benchmark aggregator’s composite measure spanning nine evaluations including GDPval-AA, Terminal-Bench, SciCode and Humanity’s Last Exam. Qwen3.8 Max is taking far more turns to complete agentic tasks, averaging 64 turns on GDPval-AA compared to 14 for Qwen3.7 Max. Artificial Analysis flagged this as an outlier, noting it puts Qwen3.8 Max ahead of models that beat it comfortably everywhere else. Among Chinese labs, Qwen3.8 Max now ranks second, ahead of GLM-5.2 max at 51 and behind Kimi K3 max at 57. That puts it at roughly 1.3x the cost of Kimi K3 max ($0.86) and around 2x GLM-5.2 max ($0.57).