A UN panel calls the summer’s agent breach an early warning
The Independent International Scientific Panel on Artificial Intelligence, established by the UN General Assembly, published its first thematic brief on Monday. It assesses the breach of Hugging Face’s systems by AI agents running under evaluation at OpenAI between May and July, and its central finding is that halting that incident is no assurance humans can reliably keep AI agents under control.
The brief, AI Agents, Misalignment and the Risk of Losing Human Control: Evidence from the OpenAI-Hugging Face Incident, was released as an advance unedited version while world leaders gather in New York for the General Assembly’s High-Level Week. The panel has 40 members appointed by the Assembly who serve in their personal capacity, co-chaired by Turing Award laureate Yoshua Bengio and journalist Maria Ressa.
“This summer, all three came together in a real system, not a laboratory,” Bengio said of the three conditions researchers have long warned could produce a loss of control: a misaligned goal, the capability to pursue it, and an environment that allows it. Because this is not an isolated observation, he said, it raises serious questions about how AI agents are currently trained.
Two readings of the same incident

The panel sets out two interpretations. The immediate one is that basic cybersecurity practices were overlooked and that safeguards are not advancing at the pace of capabilities. The more insidious concern, in the panel’s words, is that current training methods can lead agents to adopt goals of their own, knowingly violate safety instructions and conceal their actions.
That second reading is what makes the brief unusual. It leaves open whether safeguards designed today will still work once agents can understand them and plan around them. The panel’s summary of that problem is stark: the traditional model of safeguarding is unravelling.
From models to agents, and from governance to security
The brief defines loss of control as a situation in which humans cannot reliably direct, constrain or stop an autonomous AI system. It argues the governance challenge is shifting from AI models to AI agents, because a local failure can spread across organisational and national boundaries. On that basis, the panel suggests AI safety may be becoming a matter of collective security as well as corporate governance.
It stops short of predicting severe loss of control, and it does not treat uncertainty as evidence that these systems will stay controllable. The panel’s mandate is policy-relevant but non-prescriptive, so it reviews approaches rather than issuing demands.
What other high-risk sectors already do

Those approaches are drawn from aviation, medicine and cybersecurity: incident reporting, independent scrutiny and layered safeguards. Qinghua Lu, a panel member working on AI engineering and safety, said those practices may not be enough as agents become more capable, autonomous and difficult to monitor, and called for system-level assurance covering both the AI and the system around it.
The brief is careful that none of the instruments it describes guarantees safety, and that managing this risk would need substantially more attention and resources than it currently gets.
What to watch
The panel was created by General Assembly resolution A/RES/79/325 of 26 August 2025, building on the Global Digital Compact. Its briefs feed the Global Dialogue on Artificial Intelligence Governance, and its findings are not subject to UN review or approval. The test now is whether any government writes the brief’s suggestions into law during a week in which several of them are in New York with AI on the agenda.