Rankinggdp.pdf · not specified

gdp.pdf · not specified

Data updated 30 Sept 2026

Bucket
Agentic
Unit
percent
Direction
Higher is better
Version
not specified
Display harness
GPT-5.6: Frontier intelligence that scales with your ambition
Board
https://openai.com/index/gpt-5-6/

Compare published benchmark results with category weights →

Models

Published configurations retain their source and harness labels. Missing results remain unknown.

26–27 of 27 entries

ModelScoreHarnessEvidenceSource-recorded date
GPT-6.1 SolOpenAI31.8%
Reported settings & source

Reported reasoning effort xhigh. Professional questions about complex PDFs across ten professional domains. OpenAI research environment or API; same named metric and evaluation chart. No fallback is reported for this OpenAI configuration.

Provider-published lab self-report. Published chart cost per task $0.3681; source-specific cost does not establish consensus workload efficiency. Competitor results are republished from public reports; the page does not establish independent reproduction by OpenAI.

Introducing GPT-6.1 Sol · Chart: GDP.pdf / GPT-6.1 Sol / xhigh · reviewed 2026-09-29

Introducing GPT-6.1 Sollab self-report2026-09-29
GPT-6.1 SolOpenAI31%
Reported settings & source

Reported reasoning effort max. Professional questions about complex PDFs across ten professional domains. OpenAI research environment or API; same named metric and evaluation chart. No fallback is reported for this OpenAI configuration.

Provider-published lab self-report. Published chart cost per task $0.4199; source-specific cost does not establish consensus workload efficiency. Competitor results are republished from public reports; the page does not establish independent reproduction by OpenAI.

Introducing GPT-6.1 Sol · Chart: GDP.pdf / GPT-6.1 Sol / max · reviewed 2026-09-29

Introducing GPT-6.1 Sollab self-report2026-09-29