Claude Opus 5(自适应推理、最大努力模式)在Artificial Analysis Intelligence Leaderboard上获得第一名 1。该模型在满分170个模型评估中得分61分 1。
Artificial Analysis Intelligence Leaderboard对586个模型进行了多维度对比评估 1,涵盖智能程度、定价、输出速度、延迟和上下文窗口等性能指标 1。其中开源模型占94个 1。
排行前五分别为:Claude Opus 5(61分)、Claude Opus 5 Xhigh Effort(60分)、Claude Fable 5(60分)、GPT-5.6 Sol(59分)和Claude Opus 5 High Effort(59分) 1。在开源模型中,GLM-5.2 (max)排名最高,得分51分 1。
在其他性能指标方面,Mercury 2实现了最快的输出速度,达901.6 tokens/秒 1;Nova Micro提供了最低价格,为每百万tokens 0.03美元 1;Gemini 2.5 Flash-Lite的首token延迟最低,为0.33秒 1。
Claude Opus 5 with Adaptive Reasoning and Max Effort mode has claimed the top position on the Artificial Analysis Intelligence Leaderboard, achieving a score of 61 points 1. The leaderboard conducted a multidimensional comparative assessment of 586 models across various performance metrics, including intelligence level, pricing, output speed, latency, and context window capabilities, with scores derived from evaluations of 170 models 1.
The top five performers on the leaderboard are Claude Opus 5 with 61 points, Claude Opus 5 Xhigh Effort with 60 points, Claude Fable 5 with 60 points, GPT-5.6 Sol with 59 points, and Claude Opus 5 High Effort with 59 points 1. Among the 170 models evaluated, 94 are open-source 1.
Beyond overall intelligence rankings, the leaderboard tracks performance across specialized categories: Mercury 2 leads in output speed at 901.6 tokens per second, Nova Micro offers the lowest pricing at $0.03 per 1 million tokens, and Gemini 2.5 Flash-Lite achieves the lowest first-token latency at 0.33 seconds 1. Among open-source models, GLM-5.2 (max) ranks highest with a score of 51 points 1.
评论
还没有评论,欢迎留下第一条。