DataLearner logo

Qwen3.5-Omni-Flash Benchmark Details

Qwen3.5-Omni-Flash currently shows benchmark results led by MMMU-Pro (156 / 229, score 64.70), Terminal Bench Hard (186 / 244, score 8.30). This page also compares it with 2 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

Qwen3.5-Omni-Flash

Benchmark Results

Thinking
Tool usage

Agentic Development

1 evaluations
Benchmark / mode
Score
Rank/total
Terminal Bench Hard
Standard ModeTools
8.30
186 / 244

Visual Understanding

1 evaluations
Benchmark / mode
Score
Rank/total
MMMU-Pro
Standard Mode
64.70
156 / 229

Competitor Comparison

Benchmark scores for Qwen3.5-Omni-Flash compared against top models in its class

Qwen3.5-Omni-FlashMiniCPM-V 4.6Gemma 4 E4B
Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

2 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkQwen3.5-Omni-FlashCurrentGemma 4 E4B
Terminal Bench Hard
Accuracy
Agentic Development
8.30Standard Mode | Tools
8.30Thinking Enabled | Tools
MMMU-Pro
Accuracy
Visual Understanding
64.70Standard Mode
52.60Thinking Level · High

Standard API Pricing: Qwen3.5-Omni-Flash vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

Comparable standard text pricing is not available for these models.

Version History

How each version of the Qwen3.5-Omni-Flash series stacks up on benchmark tests

Qwen3.5-Omni-FlashQwen3-Omni-30B-A3B
Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

2 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.

BenchmarkQwen3.5-Omni-FlashCurrentQwen3-Omni-30B-A3B
Terminal Bench Hard
Accuracy
Agentic Development
8.30Standard Mode | Tools
3.80Thinking Enabled | Tools
MMMU-Pro
Accuracy
Visual Understanding
64.70Standard Mode
60.20Thinking Enabled

Single-Benchmark Version Trend

Viewing: Terminal Bench Hard · Agentic Development

Benchmark
NormalNormal + ToolsThinkingThinking + ToolsDeepDeep + Tools

X-axis shows model and release date, Y-axis shows score; solid lines connect the same mode across versions, while dotted guides align modes within the same generation.

Standard API Pricing Across the Qwen3.5-Omni-Flash Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

Comparable standard text pricing is not available for these models.