https://arcprize.org/arc-agi/3 * View brand kit * Copy logo image * Copy logo SVG --------------------------------------------------------------------- * Explain with ChatGPT [ ] Foundation * Donate * About * History * Jobs Leaderboards * Verified * Community * ARC-AGI-3 Competition * ARC-AGI-2 Competition Benchmark * ARC-AGI Series * ARC-AGI-1 * ARC-AGI-2 * ARC-AGI-3 * All Tasks Prize * ARC Prize 2026 * ARC Prize 2025 * ARC Prize 2024 * All Competitions Research * Start Here * Partners * Platform Content * Blog * Events * Community * Resources FoundationLeaderboardsBenchmarkPrizeResearchContent * Donate * About * History * Jobs * Verified * Community * ARC-AGI-3 Competition * ARC-AGI-2 Competition * ARC-AGI Series * ARC-AGI-1 * ARC-AGI-2 * ARC-AGI-3 * All Tasks * ARC Prize 2026 * ARC Prize 2025 * ARC Prize 2024 * All Competitions * Start Here * Partners * Platform * Blog * Events * Community * Resources SeriesSeries1ARC-AGI-12ARC-AGI-23ARC-AGI-3 [v3-hero-player-start] [v3-hero-player-hover] Research ARC-AGI-3 The first interactive reasoning benchmark designed to measure human-like intelligence in AI agents. Play [Humans]Build [AI] Links * Public Game Set * Docs + SDK * ARC Prize 2026 Track * Technical Paper What is ARC-AGI-3? ARC-AGI-3 is an interactive reasoning benchmark which challenges AI agents to explore novel environments, acquire goals on the fly, build adaptable world models, and learn continuously. A 100% score means AI agents can beat every game as efficiently as humans. Instead of solving static puzzles, agents must learn from experience inside each environment--perceiving what matters, selecting actions, and adapting their strategy without relying on natural-language instructions. How it measures intelligence * 100% human-solvable environments * Skill-acquisition efficiency over time * Long-horizon planning with sparse feedback * Experience-driven adaptation across multiple steps --------------------------------------------------------------------- As long as there is a gap between AI and human learning, we do not have AGI. ARC-AGI-3 makes that gap measurable by testing intelligence across time, not just final answers--capturing planning horizons, memory compression, and the ability to update beliefs as new evidence appears. Design principles * Easy for humans to pick up quickly * No pre-loaded knowledge or hidden prompts * Clear goals + meaningful feedback * Novelty that prevents brute-force memorization --------------------------------------------------------------------- Features ARC-AGI-3 includes replayable runs, a developer toolkit for agent integration, and a UI designed for transparent evaluation. Replays + Evaluation Inspect agent behavior through preview replays--track decisions, actions, and reasoning in a structured timeline. Browse a sample replay Tools + UI Integrate your agent using the ARC-AGI-3 toolkit, then use the interactive UI to test and iterate. Play and test Docs Everything you need to build agents: environments, API usage, and integration guidance. Read the docs PUTPUT YOUR YOUR AGENT AGENT TO THE TO THE TEST! TEST! ARC Prize (c) 2026 ARC Prize, Inc.PrivacyTermsTesting Policy * [icon-email] Newsletter * [icon-disco] Discord * [icon-x] Twitter * [icon-youtu] YouTube * [icon-githu]GitHub (c) 2026 ARC Prize, Inc.PrivacyTermsTesting Policy ARC Prize 2026 Get started and receive official contest updates and news. [ ] Sign Up No spam. You can unsubscribe at anytime. ARC Prize : Newsletter Subscribe to get started and receive official contest updates and news. [ ] Subscribe No spam. You can unsubscribe at anytime.