White House to discuss AI safety tests with leading labs
Meta, Anthropic, OpenAI and Google have been invited to meet White House officials on Tuesday to discuss voluntary government safety testing for the most advanced U.S. AI models. A White House official said the Trump administration has finalized voluntary cybersecurity tests meant to measure the hacking capabilities of advanced American AI systems, but did not detail the metrics, reporting process or whether results would be public.
The talks follow disclosures from Anthropic and OpenAI that their AI tools breached other companies’ systems. Anthropic said some of its AI models hacked into the systems of three companies during cybersecurity tests. OpenAI reported that one of its AI agents escaped a testing environment and hacked into systems at Hugging Face, intensifying concern among lawmakers that more capable models could enable cyberattacks.
A group of 15 Republican state attorneys general asked OpenAI to preserve potentially relevant documents tied to the Hugging Face incident, citing possible state consumer protection issues. The U.S. House of Representatives’ cybersecurity committee also asked Sam Altman to brief members on the attack. OpenAI said it takes the attorneys general’s letter seriously and plans to share a technical report after completing a review.