Reasoning
Mathematical and scientific problem solving.
Model profile / Meta
Meta's open-weight 29.6B multimodal agentic model for local deployment, with configurable reasoning, tool use, and a documented minimum 131,072-token context window.
Four areas of evidence
Mathematical and scientific problem solving.
Software engineering and programming.
Factual reliability and instruction following.
Tool use, workflows and professional tasks.
Current index snapshot · 2026-09-16. Each category uses its own scale; scores across categories are not directly comparable. Counts describe published catalogue evidence. A point lead does not establish superiority.
Performance in context
Muse Glimmer 30B has no current Capabilities rating. Explore the rated models below.
Lumina Capabilities Index
The current Capabilities cohort, with available operational measurements.
Chart loads as you explore
Lumina Capabilities Index
The current Capabilities cohort, with available operational measurements.
Chart loads as you explore
Explore the detail
No current index estimate. Available benchmark measurements and source records are retained below.
| Benchmark | Result | Published configuration | Source |
|---|---|---|---|
| AA Long Context Reasoning100-question set | 80% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Artificial Analysis Long Context Reasoning documented evaluation | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · ranking-eligible |
| AA Long Context Reasoning100-question set | 80% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Artificial Analysis Long Context Reasoning documented evaluation | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| AIME 20262026 / 30 questions | 94.7% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta AIME 2026 evaluation averaged across ten runs | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · ranking-eligible |
| AIME 20262026 / 30 questions | 94.7% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta AIME 2026 evaluation averaged across ten runs | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| DeepSearchQA900 questions | 74.6% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta autonomous browsing evaluation averaged across four runs | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| DeepSearchQA900 questions | 74.6% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta autonomous browsing evaluation averaged across four runs | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| OmniDocBench 1.5v1.5 / 1,355 prompts | 75.8% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta modified OmniDocBench v1.5 scoring implementation, four runs | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| OmniDocBench 1.5v1.5 / 1,355 prompts | 75.8% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta modified OmniDocBench v1.5 scoring implementation, four runs | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| Vals SkillsBench86 tasks / with skills | 44.3% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta SkillsBench evaluation with mounted skill folders, four attempts | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |
| Vals SkillsBench86 tasks / with skills | 44.3% | Muse Glimmer-30B; high reasoning; temperature=1.0; top_p=0.95; top_k=64Meta SkillsBench evaluation with mounted skill folders, four attempts | Muse Glimmer Evaluation Methodology ↗Observed 2026-08-10 · checked 2026-08-10provider-reported · reference-only |