Reasoning
Mathematical and scientific problem solving.
Score and research range
- Full-precision point
- 109.29667136454586
- Range
- 106.68–111.92 · 90% Research Range
- Scope
- Conditional fuller R0 A
Score includes qualified prediction of missing evidence.
Model profile / Upstage
Upstage's open-weight 250B-A15B hybrid-attention Mixture-of-Experts model with a 1,000,000-token context window and published high-reasoning English benchmark table.
Four areas of evidence
Mathematical and scientific problem solving.
Score includes qualified prediction of missing evidence.
Software engineering and programming.
Factual reliability and instruction following.
Tool use, workflows and professional tasks.
Current index snapshot · 2026-09-16. Each category uses its own scale; scores across categories are not directly comparable. Evidence counts describe the retained Capabilities inputs; a family can contain several tasks or configurations. A point lead does not establish superiority.
Performance in context
Solar Open 2 250B is highlighted wherever a compatible measurement is available.
Solar Open 2 250B: No current checkpoint-bound first-party quote qualified from reviewed original sources; retained values are historical only
Lumina Capabilities Index
The current Capabilities cohort, with this model highlighted.
Chart loads as you explore
Solar Open 2 250B: No exact base-profile release and documented 10K output measurement; other effort/protocol/alias records not substituted
Lumina Capabilities Index
The current Capabilities cohort, with this model highlighted.
Chart loads as you explore
Explore the detail
Score uses the admitted direct native evidence under the selected method.
Score includes qualified prediction of missing evidence.
| Benchmark | Result | Published configuration | Source |
|---|---|---|---|
| AA Long Context Reasoningstandard | 62.3% | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · ranking-eligible |
| AIME 20262026 | 95.7% | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · ranking-eligible |
| APEX-Agents-AAstandard | 16.6% | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · ranking-eligible |
| Cultural and Linguistic Intelligence in Koreanstandard | 90.7% | Solar Open 2 250B (high)Solar Open 2 card Korean table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · reference-only |
| HAE-RAE Math 8Kstandard | 92.2% | Solar Open 2 250B (high)Solar Open 2 card Korean table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · reference-only |
| KMMLU-Prostandard | 78.4% | Solar Open 2 250B (high)Solar Open 2 card Korean table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · reference-only |
| GDPval-AA v2 (Elo)v2 | 1,128 Elo | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · reference-only |
| GPQA DiamondDiamond | 86.3% | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · ranking-eligible |
| HMMT Feb 20262602 | 93.9% | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · ranking-eligible |
| Humanity's Last Examwithout tools | 28.8% | Solar Open 2 250B (high)Solar Open 2 card English table | Solar Open 2 250B model card ↗Observed 2026-08-15 · checked 2026-08-15provider-reported · ranking-eligible |
No specialist task observations retained for this model.