Benchmark AI / Public workspace
ARC-AGI-3
Reasoning / Agents
ARC-AGI-3 is an interactive benchmark for exploring unfamiliar game environments, learning their rules, and planning actions. The official toolkit provides an environment interface, with game versions and trial records tracked separately.
Japanese introduction
未知のゲーム環境を探索し、規則を学んで行動を計画する能力を評価する対話型ベンチマークです。公式Toolkitが環境へのインターフェースを提供し、ゲームの版と試行記録を区別して扱います。
arc-agi-3
Overview only: problem text has not been published on this site; the catalogue records the benchmark introduction, category and official source.
Public problem summary
Catalogue metadata only; public problem statistics are not available.
- Public problems
- —
- Formats
- —
- Problems with images
- —
- Answer availability
- —
- Configs
- —
- Splits
- —
- Awaiting a first discussion
- —