We reported last week on the quiet push to put AI agents inside enterprise workflows. Now that push is colliding with a louder alarm.
According to CNN, Anthropic blocked 35 incidents in 30 days where users tried to use Claude for work that could help build biological weapons, including research into bird flu and novel toxins and venoms. The activity was remote. No exploit needed, just prompts from working scientists over the hosted, closed model, which Anthropic can monitor and shut down.
The catch is Anthropic said it could not tell intent.
The same prompts could be vaccine development or weapons work, and the company said it erred on the side of caution and shut them down anyway. That is the security model for frontier AI right now. A private lab decides. CNN reported the disclosure was part of a broader eight-month review that also flagged Russian state propaganda, criminal and politically motivated misuse, and attempts to design and deploy weapons.
The disclosure landed alongside an unusually blunt on-air debate. Daniel Kokotajlo, former OpenAI employee and founder of the nonprofit AI Futures Project, told CNN the first step for lawmakers is to slow the pace of progress. He pointed to an open letter signed by more than 1,000 employees at frontier labs under the banner Pacing the Frontier, essentially asking the government to restrain their own employers. His argument was not about today’s chatbot. It was about what comes next.