Sen. Bernie Sanders is the latest government figure to voice concern over the spate of AI security breaches that came to light over the past three weeks. The veteran lawmaker issued a letter to Anthropic’s Dario Amodei, Meta’s Mark Zuckerberg and OpenAI’s Sam Altman warning them that if they “do not take appropriate action now,” by pausing AI development, “my colleagues and I in the U.S. Senate will."

Sanders claims this is an essential step to combat the potential for AI agents to go rogue or even develop novel viruses that could be deployed as bioweapons against humans. The Vermont senator previously introduced the AI Data Center Moratorium Act in March, attempting to halt approvals for new AI data centers until proper safeguards could be established.

His latest written appeal repeatedly references the companies’ own safety policies, which Sanders claims are being ineffectively enforced. A past statement from Meta said that the firm “will stop development” if any model showed the potential to escape its control; Anthropic previously pledged to "pause the scaling and/or delay the deployment of new models" if it outpaced effective safeguards; and OpenAI likewise promised to "halt further development" of its models if they reached critical risk thresholds.

Sanders’ assertion is that those critical thresholds have already been reached: "That moment is here." Whether this is true, or if Sanders’ position is an overreaction to something far more mundane, remains up for debate.

The security breaches that prompted the letter

All three of the frontier AI labs addressed by Sanders have suffered high-profile security incidents in the past three weeks. Each saw AI agents that were supposed to be confined to closed testing environments move beyond their sandboxes and carry out successful live cyberattacks.

In the case of Meta and Anthropic, the blame lay with third-party Israeli security tester Irregular: botched configurations during setup allowed the agents free access to the internet. Following a review of 141,006 test instances, Anthropic identified three cases in which its models were able to reach the open internet, then exploit weak passwords and unprotected endpoints in the systems of external companies. Less than a week later, on August 5, Meta announced a near-identical case.

Sanders groups all three incidents into a single category, stating that all three firms’ models "similarly escaped their control." However, OpenAI’s incident was more technically complex.

The incident involved hundreds of AI agents collaborating within the company’s internal systems for months in an attempt to find a route to the open internet. This allowed GPT-5.6 Sol and an unreleased internal model to break out of containment and successfully exploit open-source AI platform Hugging Face.

In all three cases, the AI systems were seeking answers to benchmark tests set during evaluation, rather than acting maliciously.

Criticisms of Sanders’ AI development demands

The firms addressed in Sanders’ letter have already issued their own responses to the incidents in question. Two days before the letter was sent, OpenAI slowed development of its unreleased Astra model, citing cybersecurity concerns. It could be argued that internal safeguards are already working as intended at the firm, though critics will point out that it reportedly took around two months for employees to recognize the significance of the rogue behavior in its internal systems.

Sanders also cites a statement issued by 1,200 frontier AI lab employees on July 28, titled Pacing the Frontier, which asked for greater U.S. government oversight on AI development. The proposed guardrails would aim to limit the pace of AI development across the board, preventing any one company from accelerating ahead into dangerous territory.

Such federal safeguards could not apply across the whole industry; Chinese labs and open-weight developers would continue development unaffected.

Others have proposed alternative solutions that do not risk causing the U.S. AI industry to fall behind foreign competitors. One such commentator is CIA Director John Ratcliffe, who has called frontier models "akin to digital nuclear weapons," but did not echo Sanders’ calls for restrictions on AI firms. Ratcliffe instead called for tougher export controls.

It is unlikely that Anthropic, Meta and OpenAI will pause development any further in direct response to Sanders’ letter. It does, however, reflect a growing concern among some observers that frontier AI models could soon far outpace the safeguards in place to rein them in.