NVIDIA AVO reaches 100% on ARC-AGI-3 benchmark

NVIDIA AVO completed all 183 ARC-AGI-3 levels across 25 public environments, figuring out tasks with no instructions, rules, or stated goals. Claude Opus 5 alone scored 30% on the benchmark; wrapped in AVO's harness it hit 100%. AVO is NVIDIA's general-purpose agent architecture for long-horizon autonomous work.
5 sources
NVIDIA AVO Reaches 100% on ARC-AGI-3, Demonstrating a Frontier-Level General-Purpose Architecture for Long-Horizon Autonomous Agentsdeveloper.nvidia.com
Claude Opus 5 scored 30% on ARC-AGI-3. Wrapped in Nvidia’s AVO, it hit 100%.thenewstack.io
NVIDIA’s coding agent scored 100% on ARC-AGI-3 interactive reasoning benchmarkreddit.com
NVIDIA AVO got 100% on ARC-AGI-3. It completed all 183 levels across all 25 public environments, figuring out what to do with no instructions, explicit rules, or stated goals.xcancel.com
Nvidia AVO scores 100% on the ARC-AGI-3 interactive reasoning benchmarktwitter.com
NVIDIA by email
Get an email when NVIDIA has news
No news that day, no email.
More stories today
- Reduce RAG costs on Amazon Bedrock with query-aware compression
- AWS and Panasonic Avionics use agentic AI for aircraft IFEC diagnostics
- User's Claude memory system backfired; Claude said user was the bottleneck
- Claude users discuss when to choose Sonnet over Opus
- Building token-efficient multi-agent systems