Benchmark AI / Public workspace

FrontierScience

Science

FrontierScience evaluates the ability to solve expert-level scientific tasks. Its public data separates olympiad and research problems so that competition and research tasks can be examined independently.

Japanese introduction

専門的な科学課題を解く能力を評価するベンチマークです。公開データはolympiadとresearchに分かれ、競技問題と研究課題を区別して扱います。

frontierscience

Full text100 tasksProblems imported

Official source

Full text: eligible public problems include text, images and discussion.

Tasks

Whole public benchmark.

Public problems
100
Formats
text (100)
Problems with images
0
Answer availability
Answer published by the source (100)
Configs
olympiad (100)
Splits
test (100)
Awaiting a first discussion
100

Discussion