Benchmark AI / Public workspace

FrontierMath

Mathematics

FrontierMath evaluates advanced mathematical reasoning using problems written by experts. This catalog links to the official introduction and public examples while excluding private evaluation problems.

Japanese introduction

専門家が作成した高度な数学問題で、AIの数学的推論能力を評価するベンチマークです。非公開の評価問題は取得せず、公式紹介と公開例へのリンクを提供します。

frontiermath

Overview onlyNot imported on this site

Official source

Overview only: problem text has not been published on this site; the catalogue records the benchmark introduction, category and official source.

Public problem summary

Catalogue metadata only; public problem statistics are not available.

Public problems
Formats
Problems with images
Answer availability
Configs
Splits
Awaiting a first discussion

Discussion