Reasoning
Mathematical and scientific problem solving.
Model profile / Cognition
SWE-1.7 is a proprietary model variant recorded in the BenchLM public dataset.
Four areas of evidence
Mathematical and scientific problem solving.
Software engineering and programming.
Factual reliability and instruction following.
Tool use, workflows and professional tasks.
Current index snapshot · 2026-09-16. Each category uses its own scale; scores across categories are not directly comparable. Counts describe published catalogue evidence. A point lead does not establish superiority.
Performance in context
SWE-1.7 has no current Capabilities rating. Explore the rated models below.
Lumina Capabilities Index
The current Capabilities cohort, with available operational measurements.
Chart loads as you explore
Lumina Capabilities Index
The current Capabilities cohort, with available operational measurements.
Chart loads as you explore
Explore the detail
No current index estimate. Available benchmark measurements and source records are retained below.
| Benchmark | Result | Published configuration | Source |
|---|---|---|---|
| FrontierCode 1.1 Main2026 | 42.3% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 21 July 2026 ↗Observed 2026-07-21 · checked 2026-07-21source-checked · reference-only |
| FrontierCode 1.1 Main2026 | 42.3% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 2026-07-27 ↗Observed 2026-07-27 · checked 2026-07-27source-checked · reference-only |
| FrontierCode 1.1 Main2026 | 42.3% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 2026-08-01 ↗Observed 2026-08-01 · checked 2026-08-01source-checked · reference-only |
| SWE-bench Multilingual2025 | 77.8% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 21 July 2026 ↗Observed 2026-07-21 · checked 2026-07-21source-checked · reference-only |
| SWE-bench Multilingual2025 | 77.8% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 2026-07-27 ↗Observed 2026-07-27 · checked 2026-07-27source-checked · reference-only |
| SWE-bench Multilingual2025 | 77.8% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 2026-08-01 ↗Observed 2026-08-01 · checked 2026-08-01source-checked · reference-only |
| Terminal-Bench2026 | 81.5% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 21 July 2026 ↗Observed 2026-07-21 · checked 2026-07-21source-checked · reference-only |
| Terminal-Bench2026 | 81.5% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 2026-07-27 ↗Observed 2026-07-27 · checked 2026-07-27source-checked · reference-only |
| Terminal-Bench 2.02026 | 81.5% | Exact BenchLM registry variant SWE-1.7; bulk export does not retain a complete upstream harness configuration.Source-native system | BenchLM public datasets — 2026-08-01 ↗Observed 2026-08-01 · checked 2026-08-01source-checked · reference-only |
| Benchmark | Result | Execution & attribution | Original source |
|---|---|---|---|
| FrontierCode 1.1 MainFrontierCode 1.1 Main · Cognition release table | 42percent (published) | Release-table configuration; per-model effort and harness vary. See original footnotes.provider subjectSource-native comparison only. Different settings are not silently substituted into retained index protocols. | Cognition ↗Published 2026-09-10 · checked 2026-09-16 |
| DeepSWE 1.1DeepSWE 1.1 · Cognition release table | 37.7percent (published) | Release-table configuration; per-model effort and harness vary. See original footnotes.provider subjectSource-native comparison only. Different settings are not silently substituted into retained index protocols. | Cognition ↗Published 2026-09-10 · checked 2026-09-16 |
| Terminal-Bench 2.1Terminal-Bench 2.1 · Cognition release table | 81.5percent (published) | Release-table configuration; per-model effort and harness vary. See original footnotes.provider subjectSource-native comparison only. Different settings are not silently substituted into retained index protocols. | Cognition ↗Published 2026-09-10 · checked 2026-09-16 |
| Terminal-Bench 4Terminal-Bench 4 · Cognition release table | 7.6percent (published) | Release-table configuration; per-model effort and harness vary. See original footnotes.provider subjectSource-native comparison only. Different settings are not silently substituted into retained index protocols. | Cognition ↗Published 2026-09-10 · checked 2026-09-16 |
No specialist task observations retained for this model.