According to Artificial Analysis tests, Claude Opus 5 ranks first on the comprehensive intelligence index with a score of 61, ahead of Fable 5 (60) and GPT-5.6 Sol (59). The index integrates nine tests covering knowledge work, coding, scientific reasoning, and factual accuracy. In scientific reasoning, Opus 5 scores 53% on the 'Humanity's Last Exam,' tying with Fable 5; but on the physics benchmark CritPt, it trails GPT-5.6 Sol, GPT-5.5 Pro, and GPT-5.6 Terra.
Independent tests from Epoch AI yield similar but more conservative conclusions: Opus 5's overall capability index is 159, slightly below Fable 5's 161; however, on the software engineering index, both score 161, tying for the lead. This indicates that competition among frontier models is extremely tight, with no model establishing a clear lead.