TAU2
Measured Jul 10, 2026Source
Score
0.63
OpenAI o1 is a reasoning model designed to solve complex problems in science, coding, and mathematics. It employs an internal chain-of-thought process to analyze questions before responding, significantly improving accuracy on tasks requiring deep logical inference.
Benchmark history
Score
0.63
Score
0.13
Score
0.59
Score
0.7
Score
0.72
Score
0.97
Score
0.36
Score
0.68
Score
0.08
Score
0.75
Score
0.84
Score
39.7
Score
23.4
Plan availability
Loading ratings...

Thinking... Make sure you are connected to GitHub server