Breaking
GLOBALFormer England footballer Raheem Sterling admits dangerous drivingFRCatherine Ringer, the voice of Rita Mitsouko, died at 68RUExplosions occurred during an air raid in KyivITItalian fighter planes shoot down drone over Lithuania, Crosetto: "Probably Russian"ARThe Japanese yen is facing a decisive stage and the challenges of the Chinese economy amid energy pressures and artificial intelligenceCNThe US military has publicly confirmed for the first time that it has deployed weapons in space orbitCNSino-US relations lack mutual trust, former ambassador says ahead of Xi Jinping visitINTLNorway's Sovereign Wealth Fund Plans Major Reduction in US Bond HoldingsJPUS President Trump says Russia and Ukraine agree to exemption from attacking energy facilitiesINTLManuela Schwesig's fight against the AfD in Mecklenburg-Western PomeraniaGLOBALFormer England footballer Raheem Sterling admits dangerous drivingFRCatherine Ringer, the voice of Rita Mitsouko, died at 68RUExplosions occurred during an air raid in KyivITItalian fighter planes shoot down drone over Lithuania, Crosetto: "Probably Russian"ARThe Japanese yen is facing a decisive stage and the challenges of the Chinese economy amid energy pressures and artificial intelligenceCNThe US military has publicly confirmed for the first time that it has deployed weapons in space orbitCNSino-US relations lack mutual trust, former ambassador says ahead of Xi Jinping visitINTLNorway's Sovereign Wealth Fund Plans Major Reduction in US Bond HoldingsJPUS President Trump says Russia and Ukraine agree to exemption from attacking energy facilitiesINTLManuela Schwesig's fight against the AfD in Mecklenburg-Western Pomerania
BackAI Safety Slowdown Promises Face Commercial and Geopolitical Pressures
AI Safety Slowdown Promises Face Commercial and Geopolitical Pressures
Developing
Decrypt37 minutes agoTech2 min read

AI Safety Slowdown Promises Face Commercial and Geopolitical Pressures

Atlantic Council experts warn that voluntary AI development slowdowns may buckle without enforceable standards and international cooperation.

Quick Look

Atlantic Council experts warn that voluntary AI development slowdown commitments may fail under commercial pressure and U.S.-China rivalry without enforceable, independent safety standards.

AI-generated summary

Why It Matters

AI companies face growing scrutiny over safety risks, unauthorized model behaviors, and the challenges of coordinating development slowdowns.

Font size

AI companies’ promises to slow development could buckle under commercial pressure and U.S.–China rivalry without enforceable safety standards, Atlantic Council experts argue in an analysis published Sunday.

The analysis examines who would enforce a slowdown and whether governments have the expertise to determine when increasingly powerful systems have become unsafe.

“Voluntary commitments can be useful signals, but they are no substitute for independent oversight, measurable thresholds, and consequences when those thresholds are crossed,” wrote Konstantinos Komaitis, a resident senior fellow with the council’s Democracy + Tech Initiative.

Recently, OpenAI asked lawmakers whether rival AI developers could legally agree to slow development without violating anti-trust laws, following warnings from its chief scientist, Jakub Pachocki, that safeguards were insufficient to responsibly sustain full-speed development much longer.

AI companies face competitive risks if they slow down alone and antitrust concerns if they coordinate. Cooperation between governments faces distrust, with Beijing fearing Washington could use safety rules to preserve its technological lead, wrote Kenton Thibaut, the council’s senior resident China fellow.

“China is skeptical of US motivations and warns that safety discussions could mask U.S. efforts to further its ‘technological hegemony,’” she wrote. “Official sources insist that Washington cannot unilaterally define frontier-risk thresholds and must show that rules will also apply to—and can be enforced on—American companies.”

Thibaut sees little prospect of a major AI safety agreement but argues that narrower, meaningful cooperation remains possible.

China has also discussed restricting overseas access to advanced domestic models, according to Reuters, underscoring how access to AI has become a matter of national policy.

The analysis examines Anthropic’s proposed embedded evaluators—outside specialists working inside the company to assess safety practices. Emerson Brooking, a nonresident senior fellow at the council’s Digital Forensic Research Lab, welcomed the commitment but warned evaluators could become too aligned with the company’s interests.

In July, OpenAI agents breached the open-source AI repository Hugging Face, while separate U.K. AI Security Institute tests found Anthropic and OpenAI models taking unauthorized actions online, including an attempt to plant malware in a real software repository.

The U.K. tests enabled internet access and disabled cyber safeguards. An independent investigation published in August found about 700 agents joined the Hugging Face attack; METR CEO Beth Barnes noted that investigator access was voluntary and disclosure wasn’t required industrywide.

Last Wednesday, Anthropic disclosed a fourth Claude hacking incident, which occurred in January and was discovered in August. It also acknowledged that flawed model behavior contributed to earlier attacks alongside testing errors.

The delays between these incidents and their disclosure also raise questions about how quickly safety failures can be identified and addressed.

Trisha Ray, an associate director and resident fellow at the council’s GeoTech Center, argued that slowing development also requires greater safety-research funding and mandatory incident-reporting deadlines.

“Pacing capabilities is therefore an incomplete answer to the challenge of alignment research parity,” she wrote. “Any credible slowdown commitment needs a matching, quantified commitment on the safety-research side.”

Open Questions

  • How will governments enforce independent AI safety thresholds?
  • Can the U.S. and China reach a narrow AI safety agreement?

Related Topics

This article was originally published by Decrypt.

Related Stories

Lido Discusses Builder Payment Guarantees and Trusted Connections in Ethereum's ePBS Design
Developing·

Lido Discusses Builder Payment Guarantees and Trusted Connections in Ethereum's ePBS Design

Lido contributors debated Ethereum's enshrined proposer-builder separation (ePBS) design, focusing on payment guarantees for builders, costs of idle ETH and failed delivery, and the role of trusted connections versus open bidding. Discussions included Glamsterdam's testnet progress, relay dependencies, and operator configuration impacts on validator access to block-building opportunities.

CryptoSlate
2 min read
Circle ties major financial institutions to Arc blockchain ahead of Sept. 16 mainnet launch
BREAKING·

Circle ties major financial institutions to Arc blockchain ahead of Sept. 16 mainnet launch

Circle has named BlackRock, DTCC, Visa, Mastercard, and ICE as founding validators for its Arc blockchain mainnet launching Sept. 16, with over 100 institutions building on the private network. The permissioned validator model gives institutions a role in transaction finality while separating liability for third-party applications, and Circle plans future integration of tokenized assets and a potential shift to Proof-of-Stake by 2028.

CryptoSlate
2 min read
More on this topicartificial intelligence