Nvidia just showed that the harness, not the AI model, is now the real hero
TechCrunch 2 weeks ago 46 ● 4 sources
Nvidia published research showing that a custom “harness” with better memory handling and a supervisory component enabled Claude Opus 5 to perform long-horizon interactive reasoning tasks much more reliably. It reports a 100% score on ARC-AGI-3 with the harness, versus 30% without it. The takeaway is that agent performance—and likely cost and safety—depends more on harness design than on the underlying AI model alone.