TAU2
Measured Jul 10, 2026Source
Score
0.92
Claude Opus 4.6 is Anthropic's flagship model, designed for maximum performance on complex reasoning, coding, and analysis tasks. It features adaptive reasoning capabilities to tackle challenging problems with high accuracy and depth.
Benchmark history
Score
0.92
Score
0.46
Score
0.71
Score
0.53
Score
0.52
Score
0.37
Score
0.9
Score
43.7
Plan availability
Loading ratings...

Thinking... Make sure you are connected to GitHub server