Amazon Bedrock
Claude Sonnet 5.5 debuted at No. 2 on the Artificial Analysis Intelligence Index with 56 points, just two behind Anthropic's own Opus 5.5 at 58 and ahead of OpenAI's GPT-6 Astra at 53 and GPT-6 Sol at 48. The 18-point gain over Sonnet 5 (38 points) is the largest single-generation jump on the index. On Terminal-Bench 4.0, Sonnet 5.5 reaches 64%, ahead of both Opus 5.5 and GPT-6 Astra at 60%, a 50-point gain over Sonnet 5. On Terminal-Bench-Science, Sonnet 5.5 scores 53%, trailing only GPT-6 Astra and Opus 5.5. For classic knowledge work it essentially matches Opus 5.5 on GDPval-AA (1,844 vs 1,846 Elo), AA-Briefcase (1,811 vs 1,822), and AutomationBench-AA (71% vs 70%). On AA-Omniscience factual knowledge it answers 54% correctly versus 66% for Opus 5.5, but hallucinates less (47% vs 59%), reflecting a deliberate speed-and-efficiency trade-off.