DataLearner logo

Muse Glimmer-30B Benchmark Details

Muse Glimmer-30B currently shows benchmark results led by AA-LCR (1 / 16, score 80), IF Bench (4 / 31, score 77), AIME 2026 (6 / 19, score 94.70).

Benchmark Results

Muse Glimmer-30B

Benchmark Results

Thinking
Tool usage
Internet

General Knowledge

2 evaluations
Benchmark / mode
Score
Rank/total
83.50
63 / 188
HLE
High
22
109 / 173

Coding and Software Engineer

3 evaluations
Benchmark / mode
Score
Rank/total
76
37 / 113
51.20
39 / 55
43.60
2 / 2

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
77
4 / 31

Long Context

1 evaluations
Benchmark / mode
Score
Rank/total
AA-LCR
High
80
1 / 16

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
94.70
6 / 19

AI Agent - Tool Usage

3 evaluations
Benchmark / mode
Score
Rank/total
MCP-Atlas
HighTools
75.50
12 / 28
65.90
19 / 25
51.70
29 / 29

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
HighTools
953
6 / 6

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
DeepSearchQA
HighToolsInternet
74.60
2 / 2

Multimodal Understanding

3 evaluations
Benchmark / mode
Score
Rank/total
78.80
9 / 9
75.80
2 / 2
74
5 / 7

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
τ³-Banking
HighTools
23.50
2 / 2