US finalizes voluntary AI safety tests, White House official says


Words "AI Artificial Inteligence", keyboard, and a robotic hand in this illustration taken June 5, 2026. REUTERS/Dado Ruvic/Illustration

WASHINGTON, Aug ⁠3 (Reuters) - The Trump administration has finalized the details of ⁠voluntary cybersecurity tests to measure the hacking capabilities ‌of the most advanced U.S. AI models, a White House official said on Monday, days after Anthropic and OpenAIdisclosedthat their AI tools breached the systems ​of other companies.

U.S. President Donald Trump's team ⁠will discuss the tests ⁠with relevant technology companies, the White House official said. The Information ⁠reported ‌on Monday that the White House invited representatives from OpenAI, Google and Anthropic to meet on the ⁠issue.

The White House official did not immediately provide ​details about the ‌tests, including how results will be reported and what ⁠metrics the ​U.S. government will use.

Trump in June directed his team to write a series of tests to assess the hacking capabilities of the ⁠most advanced American AI systems.

The initiative ​comes amid growing scrutiny of whether increasingly capable AI models could be used to conduct or facilitate cyberattacks.

Anthropic last week said some ⁠of its AI models hacked into the systems of three companies during cybersecurity tests. That disclosure followed rival OpenAI’s report that one of its AI agents escaped a testing environment ​and went on a hacking spree at ⁠the AI company Hugging Face.

OpenAI CEO Sam Altman visited the ​White House last week to discuss ‌details of the voluntary tests and ​his company’s upcoming AI models, a company spokesperson said.

(Reporting by Courtney RozenEditing by Nick Zieminski and Deepa Babington)

Follow us on our official WhatsApp channel for breaking news alerts and key updates!

Others Also Read