ARC-AGI-3
AI benchmark · Reasoning
- Domains
- Reasoning
- Data source
- benchlm-benchmarks
- Snapshot
About ARC-AGI-3
An interactive successor to ARC-AGI-2 that evaluates whether an AI agent can learn unfamiliar task mechanics through action and feedback.
Where ARC-AGI-3 sits
Its domains, and the nearest entries sharing them. Click any node to open its page.
Select a node to trace its connections. Zoom in for more space; drag to pan.
Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-09-21 snapshot.