ARC Prize 2025 results point to refinement loops as the central AI reasoning trend
Original titleARC Prize 2025 Results and Analysis
AISummary
ARC Prize reports that the top Kaggle entry reached 24% on the ARC-AGI-2 private dataset at $0.20 per task, and that all winning solutions and papers are open source.
The top verified commercial model, Opus 4.5 (Thinking, 64k), scored 37.6% at $2.20 per task, while a Poetiq refinement on Gemini 3 Pro reached 54% at $30 per task.
The author argues that refinement loops are the main driver of 2025 progress, and says ARC-AGI-3 is planned for early 2026.
AIWhy it matters
The post links 2025 competition results to a broader argument about refinement loops, showing how benchmark outcomes are being read as evidence of AI reasoning progress.
Source: ARC Prize · arcprize.orgPublished · added here