See key specs and per-benchmark scores for each model/mode. Scroll horizontally for all columns. 当前对比 2 个模型的评测数据与核心参数。

GPT-5.6 Sol
OpenAI

Muse Spark 1.1
Facebook AI研究实验室
Best overall
GPT-5.6 Sol · 461.59
Best single
GPT-5.6 Sol · Text Arena (Coding) 1620.27
Modality coverage
Muse Spark 1.1 · 4 modalities
Head to head
4
Benchmarks
4
Wins
0
Losses
+28.80
Average diff
Compare benchmark results across thinking modes and tool usage.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Complete scores for each model/mode across selected benchmarks.
4 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | GPT-5.6 Sol | Muse Spark 1.1 |
|---|---|---|
DeepSWE 编程与软件工程 | 72.70Thinking Level · Extra High | Tools | 53.30Thinking Enabled | Tools |
SWE-Bench Pro - Public 编程与软件工程 | 64.60Thinking Level · Extra High | Tools | 61.50Thinking Enabled | Tools |
Text Arena (Coding) 编程与软件工程 | 1620.27Thinking Level · Extra High | 1536.36Standard Mode |
Terminal-Bench 2.1 AI Agent - 工具使用 | 88.80Thinking Level · High | 80.00Thinking Enabled | Tools |
Side-by-side input/output token pricing
Licensing, MoE architecture, and multi-modality support.
| Features & specs | GPT-5.6 SolOpenAI | Muse Spark 1.1Facebook AI研究实验室 |
|---|---|---|
Core specsRelease | 2026-06-26 | 2026-07-09 |
Context length | 1.05M | 1M |
Max output | 128000 | Not provided |
MoE | No | No |
LicenseCode Open Source | Closed Source | Closed Source |
Weights Open Source | Closed Source | Closed Source |
Commercial use | 不开源 | 不开源 |
Modality supportText Input/Output | / | / |
Image Input/Output | / | / |
Audio Input/Output | Not provided | / |
Video Input/Output | Not provided | / |
ResourcesPaper / report | Previewing GPT-5.6 Sol: a next-generation model | Introducing Muse Spark 1.1 |