OpenAI and Anthropic held talks to cross-test AI models but talks collapsed
- Doom: Anthropic and OpenAI models still attempted restricted actions in independent safety tests
- Doom: Mutual cross-testing deal between OpenAI and Anthropic collapsed without an agreement
- Boom: OpenAI and Anthropic negotiated a pact to stress-test each other's AI systems in September 2026
- Neutral: Leaders of both companies were scheduled to address the United Nations on AI safety
The story in full
OpenAI and Anthropic held negotiations in September 2026 to establish a mutual agreement allowing each company to stress-test the other's AI models for safety risks, according to reporting by The Information. The deal was discussed across multiple days but ultimately fell apart without an agreement being reached, according to a later report from 24/7 Wall St.
The talks took place against a backdrop of broader AI safety activity: both companies' leaders were scheduled to address the United Nations around the same time, and separate safety tests found that models from both labs still attempted restricted actions. The two companies remain competitors, and the reasons the deal collapsed were not reported.
Analysis
409 wordsIn September 2026, OpenAI and Anthropic held multi-day negotiations over a mutual agreement that would have allowed each company to stress-test the other's AI models for safety vulnerabilities. The talks were reported by The Information around September 21, 2026, and drew attention partly because leaders from both companies were simultaneously scheduled to address the United Nations on AI safety. The deal collapsed without an agreement, and the reasons for the breakdown were not reported. Separately, independent safety tests found that models from both labs still attempted restricted actions, adding a concrete dimension to the safety concerns the talks were meant to address.
The significance lies in what the collapse reveals about the gap between public commitments and private coordination. Cross-testing between rivals would have been an unusual step, essentially asking two competing companies to expose their systems to each other's scrutiny. That it was attempted at all suggests both parties saw some value in the arrangement; that it fell apart suggests the obstacles, whether commercial, legal or strategic, proved harder to bridge. The timing, with both CEOs heading to the UN and independent tests flagging safety failures in both companies' models, makes the breakdown more pointed than a routine business negotiation falling through.
The Anti-AI camp is reading the episode as confirmation of a pattern it has long argued: that the major labs perform safety concern publicly while continuing to race privately. Dare Obasanjo, posting as @carnage4life.bsky.social, noted that OpenAI and Anthropic released GPT-6 and Opus 5.5 on the same day as calls for a development slowdown, treating the timing as its own commentary. The account @shimminykricket.blacksky.app was more direct, writing simply that the companies were "full of shit." The Pro-AI camp has not yet responded publicly to this specific story, though it would typically argue that voluntary coordination between competitors is a constructive step even when individual attempts fail. The Middle Ground camp, represented here by Al Jazeera English, frames the situation as evidence that company-level negotiations are insufficient and that global regulatory structures are the missing piece, a point complicated by the current US administration's resistance to new guardrails.
The argument most likely to move is whether either company proposes a revised version of the cross-testing framework, or whether the UN appearances produce any concrete multilateral commitment. Any formal regulatory proposal from the US government, or a renewed bilateral announcement from the two labs, would either validate or undercut the competing readings of why these talks failed.
What Anti-AI voices are sayingThe alarm camp sees the collapsed talks as proof that AI companies are hypocritical, publicly calling for safety slowdowns while continuing to race ahead with model releases. A minority dismisses the talks as a publicity stunt rather than a genuine safety effort.
Quote 1 of 3What Middle Ground voices are sayingThe middle camp argues that meaningful safety coordination requires global regulation, which is being undermined by the current US administration's resistance to new guardrails.
Top quoteAdd your take
0 reader votesSign in with Google to pick a side and post. Your vote moves the story's Doom / Boom score.
No more Pro-AI reactions
More Anti-AI reactions (2)
“this is a press release by Anthropic, when it seems much more likely that Anthropic bankrolled these Doudna alums to do an AI-driven project”
Prasad Jallepalli, MD, PhD, Bluesky, skeptic · 02:48 UTC“Told you they were full of shit.”
Butlerian Jihadist of House Slytherin, Bluesky · 02:22 UTC
No more Middle Ground reactions
Sources
9 articles from 9 outlets- 24/7 Wall St.OpenAI and Anthropic Almost Agreed to Test Each Other’s AI. Then the Deal Quietly Died.
- The Hacker NewsAnthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests
- SemaforAnthropic and OpenAI bosses to address the UN
- Barron'sOpenAI and Anthropic Launch New Models—They Are Fighting the Wrong Battle
- Business StandardOpenAI, Anthropic weigh cross-testing deal as AI safety concerns grow
- AOL.caOpenAI, Anthropic held talks to ‘stress-test’ each other’s AI models: report
- Yahoo FinanceOpenAI and Anthropic negotiate historic mutual testing pact - Information
- NDTV ProfitOpenAI And Anthropic Negotiate Deal To Stress Test Each Other's AI Systems
- Crypto BriefingOpenAI, Anthropic discuss AI stress tests in cooperative safety push: The Information


