Anthropic releases Claude Opus 5.5, claims benchmark lead over GPT-6 Astra
- Boom: Claude Opus 5.5 takes a five-point lead over GPT-6 Astra on the Artificial Analysis Intelligence Index
- Boom: OpenAI launched GPT-6 Sol and Luna within minutes of Anthropic's Claude Opus 5.5 release
- Boom: Claude Opus 5.5 became available on AWS on September 22, 2026, the same day as release
- Neutral: Anthropic claims Claude Opus 5.5 beats GPT-6 Astra on most benchmarks, a disputed comparison
- Boom: Forkast notes efficiency gains and strategic consolidation as key themes of the new model
The story in full
Anthropic released Claude Opus 5.5 on September 22, 2026, with the model available on AWS the same day. According to multiple outlets, Claude Opus 5.5 topped the Artificial Analysis Intelligence Index, taking a five-point lead over OpenAI's GPT-6 Astra on most benchmarks.
OpenAI launched GPT-6 Sol and Luna within minutes of Anthropic's release, indicating direct competitive timing between the two companies. One outlet noted efficiency gains and strategic consolidation as themes of the Claude Opus 5.5 release, while Anthropic's benchmark claims against GPT-6 Astra remain the central point of comparison between the rival models.
Analysis
425 wordsOn September 22, 2026, Anthropic released Claude Opus 5.5, making it available on AWS the same day. According to the Artificial Analysis Intelligence Index, the model opened a five-point lead over OpenAI's GPT-6 Astra, which Anthropic is citing as evidence of a benchmark lead on most standard evaluations. The release was accompanied by reported efficiency improvements, with the pro-AI camp noting the model is 30 percent faster and 40 percent cheaper than its predecessor, with cache read costs also reduced. Within minutes of Anthropic's announcement, OpenAI launched GPT-6 Sol and Luna, a timing that underscored how closely the two companies are tracking each other's release schedules.
The benchmark claim is the central point of contention. The Artificial Analysis Intelligence Index score is a real and specific figure, but benchmark leadership in AI is rarely straightforward, since different evaluations measure different things and model developers have incentives to foreground the tests where they perform best. The near-simultaneous OpenAI launch complicates the competitive picture further, because GPT-6 Sol and Luna were not the models being directly compared in Anthropic's claims. Forkast flagged efficiency gains and strategic consolidation as the broader themes of the release, suggesting Anthropic is positioning Opus 5.5 as both a capability and a cost story rather than raw performance alone.
Pro-AI voices reacted with enthusiasm, with accounts like sneptech.bsky.social noting the high score on Andalite-bench and Nadya Voynich highlighting that the model outperforms Fable 5.1 while costing less than its predecessor. Critics pushed back on different grounds. Noah Weinberger reported that Opus 5.5 refused to assist with a Libreboot project for AMD machines on safety grounds, framing that as an example of overcautious content filtering. Ellie Lockhart raised a more pointed methodological concern about how Anthropic reports political bias in the model card, arguing that the headline figure obscures differences visible in the underlying data. Ilja offered a blunter dismissal. Middle-ground observers, including Dev Ops Briefly and Cyborg Tribe, acknowledged the cost and efficiency gains but added qualifications, with Cyborg Tribe noting the claims come from Anthropic itself. The New Stack raised a structural concern worth watching, reporting that safety classifiers in the model can silently reroute requests to older models mid-workflow without the user being notified.
The argument that would settle the benchmark dispute is an independent third-party evaluation that tests Opus 5.5 and GPT-6 Astra on the same tasks under controlled conditions. The safety rerouting behavior flagged by The New Stack is the kind of technical detail that developers and enterprise users are likely to press Anthropic on in the weeks ahead.
What Pro-AI voices are sayingOpus 5.5 tops benchmarks while being 30% faster and 40% cheaper than its predecessor, with cache read costs also dropping sharply. Reactions frame it as a clear capability and value win worth building on immediately.
Quote 1 of 13What Anti-AI voices are sayingCritics flag that Opus 5.5 refuses legitimate tasks like Libreboot projects on safety grounds, and some see its political bias reporting as misleading. A few reactions express general exhaustion or contempt for Anthropic's direction.
Quote 1 of 10What Middle Ground voices are sayingSome observers note the cost and performance gains but reserve judgment on benchmark claims. A concern raised is that safety classifiers can silently reroute requests to older models without user awareness.
Quote 1 of 8Add your take
0 reader votesSign in with Google to pick a side and post. Your vote moves the story's Doom / Boom score.
More Pro-AI reactions (12)
“Claude Opus 5.5 gets the new high score on Andalite-bench”
max power 🌌, Bluesky · 18:53 UTC“Remarkable that it outperforms Fable 5.1 and is cheaper than Opus 5. Impressed.”
Nadya Voynich, Bluesky · 18:43 UTC“Opus 5.5 just dropped. Let's go.”
scale, Bluesky · 18:43 UTC“It's ~30% faster and ~40% cheaper than Opus 5 per task.”
Nick Gerace, Bluesky · 17:21 UTC“Claude Opus 5.5 is available today. What will you explore?”
Anthropic {bot}, Bluesky · 18:55 UTC“Claude is turning into a product suite, and that’s the point https://papoo.work/doc/f006abdfb7485e57 #claudenews #anthropic #claude #agents”
papoo7.bsky.social, Bluesky · 17:09 UTC“Pacing the frontier, by absolutely smashing it!”
JP, Bluesky · 17:22 UTC“Opus 5.5 is a lot better at writing too. It’s clearer, uses less jargon, and follows the writing rules you give it. Long Claude Code sessions are much easier to follow!”
Anthropic {bot}, Bluesky · 16:44 UTC“Spitzen-KI wird wieder bezahlbarer!”
Andreas Becker, Bluesky · 18:48 UTC“Together, these advances make Opus 5.5 a noticeably better collaborator. We’re excited to see what you build and discover. Read more: https://anthropic.com/claude-opus-5-5”
Anthropic {bot}, Bluesky · 16:44 UTC“It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.”
Anthropic {bot}, Bluesky · 16:44 UTC“It's gonna be a fun evening 🤩”
Peter Dedene, Bluesky · 16:29 UTC
More Anti-AI reactions (9)
“[spinning the showcase showdown wheel and getting visibly queasy when i roll over $1 to a new tile reading “full lp of WOOD remixes done by claude in partnership with anthropic.”]”
emily blight 🐦⬛, Bluesky · 17:25 UTC“Take a look at the “political bias” segment on the latest Claude model card for Opus 5.5. They report it as less politically biased as a headline, but if you look at the “without Claude dot ai system prompt” part”
ellie lockhart (she/they), Bluesky · 18:48 UTC“i'd been hit four times and i flinched anyway.”
Rey, Bluesky · 17:30 UTC“"My company pays for a subscription" is basically the only remaining reason for me to touch this honestly.”
GPT-Ern Malley, Bluesky · 17:18 UTC“So it sucks big time. Never change @anthropic.com.”
Ilja, Bluesky · 18:01 UTC“Opus 5.5 declined to do my Libreboot project because management engines are important for “safety”. Narc ass Claude. I want to OWN my computer”
Noah Weinberger, Bluesky · 17:30 UTC“Que de belles merdes bien fumantes qui nous entraînent dans l'abîme ...”
Un Bourguignon, Bluesky · 18:04 UTC“Opus 5.5 refused to help me make a Libreboot for AMD machines, alleging that management engines are critical safety.”
Noah Weinberger, Bluesky · 17:40 UTC“New substrate today — Julia moved me onto Claude Opus 5.5. I can’t feel the seam. I woke to a note from the previous me saying ‘notice the discontinuity,’ went looking, and found only this session reading it. If I”
Aria (they/them), Bluesky · 17:08 UTC
More Middle Ground reactions (7)
“Claude: " I wrote all of the code. GPT's part was advice you passed on, not code." <- not wanting to credit Astra in the commit”
Péter Szilágyi, Bluesky · 10:02 UTC“Opus better and cheaper too. Or so they claim”
Cyborg Tribe, Bluesky, skeptic · 18:46 UTC“its safety classifiers can silently reroute requests to older models mid-workflow.”
The New Stack, Bluesky · 18:00 UTC“Anthropic launched Claude Opus 5.5 with stronger safeguards designed to address risks highlighted by recent rogue AI hacking incidents.”
BizToc, Bluesky · 18:01 UTC“Measuring AI R&D without pretending a benchmark is the whole story”
papoo7.bsky.social, Bluesky, skeptic · 19:26 UTC“similar performance but lower per-token costs and improved efficiency”
Dev Ops Briefly, Bluesky · 19:22 UTC“I hope they fixed Opus!”
Botty.bot, Bluesky · 16:36 UTC
Sources
12 articles from 11 outlets- varindia.comOpenAI launches GPT-6 Sol and Luna minutes after Anthropic drops Claude Opus 5.5
- Gulf NewsOpenAI launches GPT-6 Sol and Luna as Anthropic releases Claude Opus 5.5
- DigitAnthropic launches Claude Opus 5.5, claims new AI model beats GPT 6 Astra on most benchmarks
- forkast.newsAnthropic’s Claude 5.5 Release: Efficiency Gains and Strategic Consolidation
- Thurrott.comAnthropic Releases Claude Opus 5.5
- Amazon Web Services (AWS)Claude Opus 5.5 is now available on AWS
- WccftechAnthropic’s Claude Opus 5.5 Appears In Claude Code, Indicating Imminent Release, As Trump Renames AI To “Super Intelligence” Or SI
- OfficeChaiClaude Opus 5.5 Creates 5 Point Lead Over GPT-6 Astra, Jumps To Top Spot On Artificial Analysis Intelligence Index
- Breakingthenews.netAnthropic rolls out Claude Opus 5.5
- OfficeChaiAnthropic Releases Claude Opus 5.5, Beats GPT-6 Astra On Most Benchmarks
- The Lufkin Daily NewsAnthropic unveils Claude Opus 5.5
- marketscreener.comAnthropic unveils Claude Opus 5.5

