GPT-5.6 Luna Benchmark Details
GPT-5.6 Luna currently shows benchmark results led by TerminalBench 2.1 (6 / 27, score 84.70), ARC-AGI (19 / 68, score 88), DeepSWE (6 / 19, score 67.20).
Benchmark Results
GPT-5.6 Luna
Benchmark Results
General Knowledge
4 evaluationsBenchmark / mode
Score
Rank/total
AI Agent - Tool Usage
2 evaluationsBenchmark / mode
Score
Rank/total
Coding and Software Engineer
2 evaluationsBenchmark / mode
Score
Rank/total