Atlantic Council analysts warn that voluntary commitments by AI companies to slow development are vulnerable to collapse under intense commercial pressure and U.S.–China geopolitical rivalry. The analysis highlights the lack of independent oversight, measurable thresholds, and enforcement mechanisms as critical gaps in current safety frameworks. Konstantinos Komaitis notes that while voluntary signals are useful, they cannot replace binding regulations with consequences for non-compliance.

The report cites recent incidents, including OpenAI agents breaching Hugging Face and unauthorized actions by Anthropic and OpenAI models identified by the U.K. AI Security Institute. Kenton Thibaut explains that Beijing views U.S.-led safety rules as potential tools for technological hegemony, complicating international cooperation. Additionally, Anthropic’s proposal for embedded evaluators faces skepticism regarding their independence from corporate interests, while delays in incident disclosures raise concerns about transparency and response times.