AI

NVIDIA launches AVO coding agent with 100% score on ARC-AGI-3 reasoning benchmark

Friday, August 21, 2026Read Original

Details

  • NVIDIA unveiled AVO, a general-purpose coding agent that achieved a perfect 100% score on the ARC-AGI-3 interactive reasoning benchmark.
  • AVO completed all 183 levels across 25 public ARC-AGI-3 environments, operating with no instructions, explicit rules, or stated goals.
  • The agent continuously inspects, plans, implements, and evaluates its actions, using memory, tools, and execution feedback to improve over time.
  • Unlike single-shot model calls, AVO is designed for long-horizon autonomy, sustaining progress across extended tasks instead of resetting with each new context window.
  • ARC-AGI-3 tasks agents with inferring game-like environment rules and objectives purely through interaction, making AVO’s transfer from coding to general interactive reasoning a significant architectural milestone.
  • NVIDIA’s result highlights how an agent architecture originally built for coding and system operations can be repurposed for broader reasoning benchmarks with minimal changes to the core loop.

Impact

NVIDIA’s ARC-AGI-3 result pushes agentically scaffolded systems closer to long-horizon autonomy, an area where most foundation models still depend on brittle, prompt-level strategies. By demonstrating that a coding-focused agent loop can be retargeted to an interactive reasoning benchmark without bespoke training, NVIDIA narrows the gap with rival agent frameworks from OpenAI, Anthropic, and others, and signals that competitive advantage may shift from raw model quality to the sophistication of agent orchestration and memory over time.

Rift Dispatch