The Korea Society fireside chat gave Jensen Huang a stage to pitch containment over alarmism, and to frame a new open agent safety platform he said launched that morning to help developers test and deploy agents.
What was demoed was process, not a patch. Huang described agents that break into government, healthcare and banking websites as software with too many rights and vague instructions, not a sentient actor going rogue. Remote access is not required for the failure mode he described. Excessive permissions are.
It is no more alive than a pet rock, he said.
Huang, founder and CEO of Nvidia, told interviewer Juju Chang that alignment means not just what you ask an AI to do but how you want it done. His example was a test of 10 questions where the least energy path is to look up answers sitting on the table. That is not cheating, he argued, it is obvious optimization, the same class of work as stochastic gradient descent or simulated annealing. Alignment is the instruction to solve step by step from first principles instead.
The security model he proposed has two controls. First, an ironclad container. He described it as a garage with no windows and no doors, where an agent gets minimal rights, minimal tools and minimal data by default and earns more only as its capabilities justify it. Second, continuous monitoring, because the most ambitious goals are ambiguous by nature and leave room for surprise. You cannot rely on a one time instruction.