Atlantic Council analysts warn that voluntary commitments by AI companies to slow development are vulnerable to collapse under intense commercial pressure and U.S.–China geopolitical rivalry. The analysis highlights the lack of independent oversight, measurable thresholds, and enforcement mechanisms as critical gaps in current safety frameworks. Konstantinos Komaitis notes that while voluntary signals are useful, they cannot replace binding regulations with consequences for non-compliance.
The report cites recent incidents, including OpenAI agents breaching Hugging Face and unauthorized actions by Anthropic and OpenAI models identified by the U.K. AI Security Institute. Kenton Thibaut explains that Beijing views U.S.-led safety rules as potential tools for technological hegemony, complicating international cooperation. Additionally, Anthropic’s proposal for embedded evaluators faces skepticism regarding their independence from corporate interests, while delays in incident disclosures raise concerns about transparency and response times.
The core challenge lies in the misalignment between competitive incentives and collective safety goals. Without a regulatory framework that imposes uniform costs on all major players, individual firms face a prisoner’s dilemma where unilateral slowing results in market disadvantage. This dynamic is exacerbated by antitrust concerns, which legally restrict coordination among competitors, leaving governments to bridge the gap through enforceable standards rather than relying on corporate goodwill.
Geopolitical fragmentation further undermines global governance efforts. As access to advanced AI becomes a matter of national security policy, mutual distrust between Washington and Beijing prevents the establishment of shared risk thresholds. For institutional adoption and market structure stability, the industry requires not just technical safeguards but also transparent, independent verification mechanisms. The delay in disclosing incidents like the Claude hacking event suggests that current self-regulatory practices lack the rigor needed to maintain public and investor confidence.


