RankingFrontierMath Erdős

FrontierMath Erdős

Data updated 6 Oct 2026

Bucket
Supporting evidence
Unit
percent
Direction
Higher is better
Version
unversioned
Display harness
FrontierMath Erdős proof scorer
Board
https://epoch.ai/benchmarks/frontiermath-erdos

FrontierMath Erdős records Lean proof accuracy on 68 conjectures. These rows are display only. They stay outside weighted Overall and outside the FrontierMath tier family budget.

Compare published benchmark results with category weights →

Models

Results are ordered by score in the display harness, followed by models with no result. Other harnesses are listed separately below. Missing results remain unknown.

1–25 of 460 entries

ModelScoreEvidenceSource-recorded date
GPT-6 AstraOpenAI3%official board2026-08-28
Claude Fable 5Anthropic0%official board2026-08-28
Claude Fable 5.1Anthropic0%official board2026-09-01
GPT-5.5OpenAI0%official board2026-08-28
GPT-5.6 SolOpenAI0%official board2026-08-28
Agnes 2.5 Pro AlphaAgnes AI———
Amazon Nova 2 LiteAmazon———
Amazon Nova 2 Pro PreviewAmazon———
Amazon Nova LiteAmazon———
Amazon Nova MicroAmazon———
Amazon Nova PremierAmazon———
Amazon Nova ProAmazon———
AutoGLM-Phone-9BZ.ai———
AutoGLM-Phone-9B-MultilingualZ.ai———
aya-101Cohere———
aya-23-35BCohere———
aya-23-8BCohere———
aya-expanse-8bCohere———
aya-vision-8bCohere———
C4Ai Aya Expanse 32BCohere———
C4Ai Aya Vision 32BCohere———
c4ai-command-r7b-arabic-02-2025Cohere———
chatglm-6bZ.ai———
chatglm2-6bZ.ai———
chatglm2-6b-32kZ.ai———