Qwen3.8 Max catches Claude Opus 4.8 but Kimi K3 still scores higher for 25 percent less
The Decoder··作者 Maximilian Schreiner
资讯摘要
Qwen3.8 Max catches Claude Opus 4.8 but Kimi K3 still scores higher for 25 percent less Alibaba's Qwen3.8 Max scores 56 on the Artificial Analysis Intelligence Index, a 10-point jump over Qwen3.7 Max (46). According to Artificial Analysis, that puts it on par with Claude Opus 4.8 and ahead of GLM-5.2 (51), but behind Kimi K3 (57), which also runs 25 percent cheaper. On GDPval-AA, a benchmark for work-related tasks, Qwen jumps 468 Elo points to 1,739, passing Kimi K3 (1,685). Only Claude Opus 5 (1,852) scores higher. The catch is how it gets there. Qwen3.

来源与参考
收录于 2026-08-07