RankingFrontierCode · 1.1 Extended

FrontierCode · 1.1 Extended

Data updated 24 Sept 2026

Bucket
Supporting evidence
Unit
percent
Direction
Higher is better
Version
1.1 Extended
Display harness
Claude Fable 5.1 and Claude Mythos 5.1 System Card
Board
https://www.anthropic.com/claude-fable-5-1-mythos-5-1-system-card

The available records have no admitted matched comparison in the capability core. Raw results remain available below.

Compare published benchmark results with category weights →

Models

Published configurations retain their source and harness labels. Missing results remain unknown.

1–4 of 4 entries

ModelScoreHarnessEvidenceSource-recorded date
Claude Fable 5Anthropic64.9%
Reported settings & source

Cognition agentic coding; composite functional and code-quality score. Fable5.1 medium effort; Fable5 xhigh.

Fable is the safeguarded deployed configuration; selected tasks may use disclosed Opus fallback, except evaluations explicitly counting safety blocks as failures.

Claude Fable 5.1 and Claude Mythos 5.1 System Card · 8.4 · reviewed 2026-09-24

Claude Fable 5.1 and Claude Mythos 5.1 System Cardlab self-report2026-09-24
Claude Fable 5.1Anthropic63.6%
Reported settings & source

Cognition agentic coding; composite functional and code-quality score. Fable5.1 medium effort; Fable5 xhigh.

Fable is the safeguarded deployed configuration; selected tasks may use disclosed Opus fallback, except evaluations explicitly counting safety blocks as failures.

Claude Fable 5.1 and Claude Mythos 5.1 System Card · 8.4 · reviewed 2026-09-24

Claude Fable 5.1 and Claude Mythos 5.1 System Cardlab self-report2026-09-24
Claude Opus 5.5Anthropic65.3%
Reported settings & source

Cognition agentic coding in Claude Code; composite functional and code-quality score; medium effort; mean@5. Highest Extended score. Cognition ran the evaluation.

Provider-published lab self-report from the Claude Opus 5.5 system card (SHA256 7311c9c6bbb16d012f1c12c7418b05949fcf7ae3e30d2c40f22050074b2a7378). Not an independent board.

Comparison limit: Provider-published lab self-report. Shown as a labeled claim and not admitted as a matched board comparison.

Claude Opus 5.5 System Card · section 8.4 · reviewed 2026-09-22

Claude Opus 5.5 System Cardlab self-report2026-09-22
Claude Opus 5.5Anthropic63.6%
Reported settings & source

Cognition agentic coding in Claude Code; composite functional and code-quality score; max effort; mean@5. Cognition ran the evaluation.

Provider-published lab self-report from the Claude Opus 5.5 system card (SHA256 7311c9c6bbb16d012f1c12c7418b05949fcf7ae3e30d2c40f22050074b2a7378). Not an independent board.

Comparison limit: Provider-published lab self-report. Shown as a labeled claim and not admitted as a matched board comparison.

Claude Opus 5.5 System Card · section 8.4 · reviewed 2026-09-22

Claude Opus 5.5 System Cardlab self-report2026-09-22