Benchmark AI / Public workspace

ARC-AGI-3

Reasoning / Agents

ARC-AGI-3 is an interactive benchmark for exploring unfamiliar game environments, learning their rules, and planning actions. The official toolkit provides an environment interface, with game versions and trial records tracked separately.

Japanese introduction

未知のゲーム環境を探索し、規則を学んで行動を計画する能力を評価する対話型ベンチマークです。公式Toolkitが環境へのインターフェースを提供し、ゲームの版と試行記録を区別して扱います。

arc-agi-3

Overview onlyNot imported on this site

Official source

Overview only: problem text has not been published on this site; the catalogue records the benchmark introduction, category and official source.

Public problem summary

Catalogue metadata only; public problem statistics are not available.

Public problems
Formats
Problems with images
Answer availability
Configs
Splits
Awaiting a first discussion

Discussion