AI Briefing
KO

ARC-AGI Leaderboard

·2026.07.25 15:31

Key point

The ARC-AGI-3 leaderboard, which measures the adaptive intelligence of AI agents, has been released.

Details

ARC-AGI has evolved into a metric for measuring the fluid intelligence of AI, and its latest version, ARC-AGI-3, evaluates an AI agent's ability to instantly adapt to new interactive environments.

The leaderboard visualizes and analyzes the following key data:

  • Reasoning Systems: Reasoning system trend lines showing performance changes (Asymptotic behavior) as inference time increases
  • Base LLMs: Performance of base models such as GPT-4.5 and Claude 3.7, performed in a single-shot manner without additional reasoning
  • Kaggle Systems: Highly efficient systems for the Kaggle challenge, operating under a strict computational budget ($50)

In particular, beyond performance alone, it analyzes the relationship between cost-per-task and performance, treating efficiency—solving problems with minimal resources—as a key metric.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.