DataLearner logo

Claude Opus 5.5 Benchmark Details

Claude Opus 5.5 currently shows benchmark results led by HLE (1 / 566, score 67.70), Terminal-Bench 4.0 (1 / 92, score 66.40), CursorBench 4.0 (1 / 46, score 57.80).

Benchmark Results

Claude Opus 5.5

Benchmark Results

Thinking

General Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
HLE
Thinking Level · MaxTools
67.70
1 / 566

AI Agent - Tool Usage

3 evaluations
Benchmark / mode
Score
Rank/total
OSWorld 2.0
Thinking Level · MaxTools
81.80
1 / 13
Terminal-Bench 4.0
Thinking Level · Extra HighTools
66.40
1 / 92
Terminal-Bench-Science 0.1
Thinking Level · MaxTools
58.70
2 / 12

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
AutomationBench
Thinking Level · MaxTools
40
13 / 23

Multimodal Understanding

1 evaluations
Benchmark / mode
Score
Rank/total
Chartography
Thinking Level · MaxTools
89
1 / 9

Coding and Software Engineer

4 evaluations
Benchmark / mode
Score
Rank/total
CursorBench 4.0
Thinking Level · MediumTools
52.50
2 / 46
CursorBench 4.0
Thinking Level · MaxTools
57.80
1 / 46
FrontierCode 1.1 Main
Thinking Level · MediumTools
54.60
1 / 9
FrontierCode 1.1 Main
Thinking Level · MaxTools
54.40
2 / 9