Holy: NVIDIA’s coding agent AVO scored 100% on ARC-AGI-3’s 25 public games, solving all 183 levels.
The agent receives no rules or stated goals. It must learn by trying things, observing the results and correcting its mistakes.
AVO succeeds by remembering what it learned and building on it over long periods instead of starting over when the model’s context resets.
Powered by Claude Opus 5, the same system previously worked autonomously for seven days optimizing GPU code.
Really really cool to see what NVIDIA achieved here