Skip to content

Opus 5 Claims Top Spot on Smart Rankings by Slim Margin, Single-task Cost 26% Lower than Fable 5

Jul 25, 11:36

According to Insight Beating monitoring, the third-party evaluation agency Artificial Analysis has released the Claude Opus 5 score. It scored 61 in the comprehensive intelligence index based on 9 tests, narrowly ahead of Fable 5 with 60 points. GPT-5.6 Sol scored 59, and Kimi K3 scored 57.

Opus 5 has an average cost per task of $2.03, which is 26% lower than Fable 5's $2.75. It topped the GDPval-AA v2 and AA-Briefcase knowledge work evaluations, and also tied for first place in the Programming Agent index when paired with Claude Code. The Terminal-Bench v2.1 score is 89%, roughly equaling GPT-5.6 Sol.

The model offers 5 levels of reasoning intensity. From low to max, the Token output differs by about 8 times, and the GDPval-AA v2 score differs by 407 Elo. Users can exchange more Tokens for higher performance or actively reduce costs.

However, the weaknesses are also apparent. Opus 5's factual knowledge still lags behind Fable 5. In the AA-Omniscience test, its illusion rate has increased to 50%, which is 14 percentage points higher than Opus 4.8. The cost-effectiveness of the low reasoning levels also remains slightly inferior to the GPT-5.6 series.

View source