Two AI Models Set to “stir government urgency”, But Will This Challenge Undo Them?
AI Explained
Mar 26, 2026
Current frontier AI models struggle significantly with the new ARC AGI 3 benchmark, which emphasizes abstract reasoning, memory, and goal setting over rote knowledge. While labs race to build automated AI researchers, performance data confirms we remain in a 'messy middle' where models act as capable drafting assistants but lack the fluid, adaptive intelligence of humans.
Key insight: Human test subjects achieve a 100% baseline on ARC AGI 3, while the top AI models currently score less than half a percent on the same task.