Alibaba Updates Qwen3.8-Max, Tops CodeArena Front-End Coding
Alibaba updated its flagship Qwen3.8-Max model on September 2, with post-training for coding and professional office tasks. The new version scored 1691 on CodeArena's WebDev benchmark, a 22-point improvement, ranking first overall ahead of models like Claude Opus5 and Kimi K3. Its average cost is $5 per million tokens, placing it on the Pareto frontier for price-performance.
Coverage timeline
机器之心新闻资讯
9月2日,阿里更新旗舰模型Qwen3.8-Max,性能较旧版本显著提升。在针对编程(Coding)和专业办公(Cowork)进行专项后训练后,Qwen3.8-Max新版本整体性能增强。在聚焦前端编程能力(WebDev)的全球权威三方榜单CodeArena中,提升22分至1691分,领先Claude Opus5、Kimi K3等一众模型,位居总榜第一。同时,CodeArena更新的模型性价比榜单(帕累托前沿)显示,Qwen3.8-Max新版本每百万Tokens综合平均仅5美元,直接“斩杀”所有价格大于5美元的其他模型。据了解,Qwen3.8-Max是阿里千问迄今最强大的大语言模型,总参数达2.4万亿,支持100万上下文Tokens;更新后的Qwen3.8-Max涌现出更强的智能体编程能力,更适合企业真实复杂任务、科研、长周期任务等。目前,Qwen3.8-Max新模型已上线千问AI平台对外提供API服务,千问办公、Qoder、千问APP均第一时间接入。
