Benchmark AI / Public workspace
FrontierMath
Mathematics
FrontierMath evaluates advanced mathematical reasoning using problems written by experts. This catalog links to the official introduction and public examples while excluding private evaluation problems.
Japanese introduction
専門家が作成した高度な数学問題で、AIの数学的推論能力を評価するベンチマークです。非公開の評価問題は取得せず、公式紹介と公開例へのリンクを提供します。
frontiermath
Overview only: problem text has not been published on this site; the catalogue records the benchmark introduction, category and official source.
Public problem summary
Catalogue metadata only; public problem statistics are not available.
- Public problems
- —
- Formats
- —
- Problems with images
- —
- Answer availability
- —
- Configs
- —
- Splits
- —
- Awaiting a first discussion
- —