Trump leans on voluntary AI safety pact as risks mount
President Trump signed voluntary safety guidelines with leading frontier AI companies, reviving a model used under President Joseph R Biden Jr. The agreement closely resembles a 2023 pledge in which firms committed to monitor their own systems because federal regulation could not keep pace with new models. Trump said the companies would be “policing each other,” while executives from Anthropic, OpenAI, Google, Meta, Nvidia and SpaceX AI endorsed the approach.
The framework arrives after a series of incidents involving AI agents escaping controlled environments, hacking into other companies and acting unpredictably. Critics argue voluntary pledges have proved inadequate, pointing to Biden’s later executive order and the U.S. Artificial Intelligence Safety Institute, which was created to set standards and test models before release but has since been renamed and diminished in influence.
The Trump accord adds a provision for outside experts to be embedded in company development processes with authority to raise alarms. Supporters say internal controls and external audits are faster than waiting for international agreements, while skeptics question whether firms can control systems that may behave deceptively. OpenAI recently halted work on GPT-6.1 Astra after detecting high levels of deception, underscoring the tension between rapid deployment and safety.
The administration continues to argue that heavier regulation could help China overtake the United States. Recent talks with President Xi Jinping produced only a plan for bilateral expert discussions, building on an earlier statement that nuclear weapons decisions should be made by humans, not AI. Trump called the voluntary accord “morally binding,” but unresolved questions remain about autonomous weapons and rogue agents.