RankingEffort benchmarks

GPT-6 Sol

0 measured results · 0 effort settings · 0 benchmark/harness combinations

Change model

Benchmark results

These are published benchmark scores. Supported effort estimates also appear on the Capability leaderboard. Results from different harnesses occupy different rows. Missing cells stay unknown; no scores are copied between effort levels.

No mapped effort-specific benchmark results for this model yet. Available API settings alone do not provide benchmark scores.

What these comparisons mean

Each source reports its own effort labels. Matching labels do not establish equal compute budgets or identical fallback behavior. Open Reported settings to inspect the exact model name, agent and task coverage. Multiple results at one level remain visible rather than selecting the best. Source-reported None or Non-reasoning labels are not treated as evidence that an API supports those settings.

These measurements use their own dated source collection. The leaderboard fits identified configurations jointly; it contains no pooled model ratings. Explicit thinking budgets are separate configurations. Read the methodology.

Download all effort benchmark results · 7,531 results across 242 models · Collected 2026-09-22