Qwen3.8 Max catches Claude Opus 4.8 but Kimi K3 still scores higher for 25 percent less
Alibaba's Qwen3.8 Max reaches a score of 56 on the Artificial Analysis Intelligence Index, marking a significant improvement over its predecessor while competing with models like Claude Opus 4.8 and Kimi K3.
Alibaba has updated its Qwen series with Qwen3.8 Max, which reportedly achieves a score of 56 on the Artificial Analysis Intelligence Index. This represents a substantial performance gain compared to the previous Qwen3.7 Max iteration, which recorded a score of 46.
The new release positions Alibaba's model competitively within the current large language model landscape. While it reportedly catches up to Anthropic's Claude Opus 4.8, industry reports suggest that Moonshot AI's Kimi K3 still holds a performance edge while costing 25 percent less.
These benchmark comparisons underscore the intense competition among frontier models regarding both capability and cost efficiency. As performance metrics converge across different providers, pricing structures are becoming increasingly vital for developers and enterprises selecting inference solutions.
This page provides an editorial summary based on publicly available information. It is not a republished article. Use the source link below for the original report.