RankingSciCode-Verified · pass@1 (n=6)
SciCode-Verified · pass@1 (n=6)
- Bucket
- Supporting evidence
- Unit
- percent
- Direction
- Higher is better
- Version
- pass@1 (n=6)
- Display harness
- Introducing Mistral Large 4 — public preview
- Board
- https://mistral.ai/news/mistral-large-4/
The available records have no admitted matched comparison in the capability core. Raw results remain available below.
Compare published benchmark results with category weights →
Models
1–1 of 1 entries
| Model | Score | Harness | Evidence | Source-recorded date |
|---|---|---|---|---|
| Mistral Large 4Mistral | 91.8%Reported settings & sourcePublic Mistral Large4 Preview. Effort, sampling and evaluation budgets are not stated. Provider-published lab_self_report; retained as launch evidence, not an independent board admission. Distinct from the generic Artificial Analysis SciCode series; version retained literally from chart title. Exact printed chart label; asset https://mistral.ai/_astro/scicode-verified-pass@1-(n_6)-alt_ZYLgs8.webp?dpl=6ac6a8353c2a68969e300fd2; asset SHA256 5010c6725092f676802d30b0d1181a8c1c873b83b62fcd440f454bf223841c43. Comparison limit: Provider publication; exact benchmark-specific configuration and independent evaluation provenance require separate review before matched-board admission. Introducing Mistral Large 4 — public preview · Capabilities deep-dive; SciCode-Verified; chart 17 (mistral-chart-17.webp) · reviewed 2026-10-07 | Introducing Mistral Large 4 — public preview | lab self-report | 2026-10-07 |