RankingMiMo-V2.6-Flash
MiMo-V2.6-Flash
3 published benchmark measures · 2 benchmark families contribute across 2 task areas. 2 capability estimates available in the full profile. See all results ↓ · Compare published benchmarks →
Model evidence summary
This profile combines published settings. It is not a runnable configuration or a leaderboard rank. Compare measured configurations →
Performance profile
Capabilities
Filled points are supported; hollow points are preliminary. Lines stop at unknown capabilities. Exact values and sources follow below.
Scores estimate outcomes against a shared reference panel; they are not accuracy percentages. Sparse or disconnected evidence cannot qualify an overall profile. Open a capability to inspect its evidence.
Reported effort · Not specified
Settings reported in this model's published benchmark results, including results outside the aggregate. Effort names are provider-specific. These are not API defaults or equal compute budgets.
- Not specified: 3 observations
Mixed settings means multiple settings occur in the evidence. Best across efforts means the source selected its best reported result across settings; it does not mean Max. Unspecified settings stay unknown. This model-summary chart combines reported settings. The leaderboard keeps identified configurations separate and excludes unknown effort.
Inspect each result and its source ↓ · Download effort evidenceScore contributions and missing evidence
2 contributing families across 2 capabilities. Fixed reference panels do not change when the catalog expands.
Capability is fitted jointly across families. Capability estimates below describe different task areas; their weighted sum is not the Capability score.
Results without reviewed compatibility or a reference match remain in the raw evidence below. Coverage counts only contributing results.
Model information & shareable badge
- Lab
- Xiaomi
- Catalog status
- active
- Availability
- Public MiMo API with prepaid pay-as-you-go or Token Plan access; MIT-licensed RL weights also available from the provider. Account and region restrictions may apply.
- Family
- MiMo-V2.6-Flash
- Released
- 2026-09-22
- Context
- 1,000,000 tokens
- API list price
- $0.14 input / $0.28 output per million tokens
- License
- mit
- Model card
- https://mimo.mi.com/models/en-US/mimo-v2.6-flash
- Default Capability family coverage
/badge/mimo-v2.6-flash.svg
Benchmark scores & sources
Original results, evaluation harnesses, and evidence behind this model.
| Benchmark | Bucket | Score | Harness | Evidence | Source-recorded date |
|---|---|---|---|---|---|
| AutomationBench-AA | Supporting evidence | 64.05168250285656% | Artificial Analysis AutomationBench-AA | official board | 2026-10-08 |
| GDPval-AA | Agentic | 1611 | Artificial Analysis GDPval-AAcontributes to capability | official board | 2026-10-08 |
| LMArena Text Arena | Human pref | 1452 | LMArena Textcontributes to capability | official board | 2026-10-02 |
| ARC-AGI-2 | Hard reasoning | — | — | — | — |
| DeepSWE v1.1 | Agentic | — | — | — | — |
| GPQA Diamond | Hard reasoning | — | — | — | — |
| Humanity's Last Exam | Hard reasoning | — | — | — | — |
| LiveCodeBench | Coding | — | — | — | — |
| MMLU-Pro | Knowledge | — | — | — | — |
| OSWorld-Verified | Agentic | — | — | — | — |
| SWE-bench Pro | Agentic | — | — | — | — |
| SWE-bench Verified | Agentic | — | — | — | — |
| Terminal-Bench 2.1 | Agentic | — | — | — | — |