Qwen3-235B-A22B-2507 Benchmark Details
Qwen3-235B-A22B-2507 currently shows benchmark results led by SimpleQA (10 / 47, score 54.30), MMLU Pro (46 / 133, score 83), GPQA Diamond (158 / 270, score 77.50). This page also tracks comparisons against 1 predecessor or same-series models. 1 source link is attached for reference.
Benchmark Results
Benchmark Results
General Knowledge
4 evaluationsVersion History
How each version of the Qwen3-235B-A22B-2507 series stacks up on benchmark tests
2 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | Qwen3-235B-A22B-2507Current | Qwen2.5-72B |
|---|---|---|
MMLU Pro 综合评估 | 83.00Standard Mode | 58.10Standard Mode |
GPQA Diamond 科学与综合推理 | 77.50Standard Mode | 45.90Standard Mode |
Single-Benchmark Version Trend
Viewing: MMLU Pro · 综合评估
Standard API Pricing Across the Qwen3-235B-A22B-2507 Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier.
These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Qwen3-235B-A22B-2507 | 阿里巴巴 | ¥0.002 / 1K tokens | ¥0.008 / 1K tokens | — |
Qwen2.5-72B | 阿里巴巴 | ¥0.004 / 1K tokens | ¥0.012 / 1K tokens | — |