INDEX 47 ▼1 todaySPLIT OF THE DAY OpenAI annual recurring revenue approaches $70 billion56 STORIES · 565 REACTIONSANTI-AI 74% · MIDDLE GROUND 16% · PRO-AI 9%LATEST AI researchers warn superintelligence extinction risk is around 50 percent
4 sources0 reactions

Anthropic reports AI agents blocked one action in 47,000 runs

32 DoomStory toneSafety warning mixed with internal deployment disclosure
4 sources · oodaloop.com · RealClearEnergy · Gulf News
  • Doom: Hackers are using AI agents to breach companies in hours, Anthropic warns
  • Doom: Anthropic CEO warns AI agent swarms could pose threats to humanity
  • Neutral: Anthropic runs roughly 30,000 AI agents internally on its own systems
  • Doom: Only one in every 47,000 agent actions is blocked by Anthropic's systems
The story in full

Anthropic disclosed that it runs approximately 30,000 AI agents on its own systems and that only one action in every 47,000 is blocked. The company also reported that hackers are using AI agents to breach companies within hours.

Anthropomorphic CEO statements addressed the risk of AI "agent swarms" posing broader threats to humanity. The figures and warnings appear to come from Anthropic itself, placing the company in the position of both deploying large-scale agent systems and raising alarms about how similar technology is being weaponized by malicious actors.

Analysis

374 words

In late September 2026, Anthropic disclosed that it operates roughly 30,000 AI agents on its own internal systems and that only one action in every 47,000 is blocked by its safety mechanisms. The company also warned that hackers are exploiting similar agent technology to breach companies within hours. A separate report published on September 23 described a Chinese hacker deploying agents built on Anthropic, DeepSeek, and Moonshot AI in a coordinated cyberattack targeting approximately 100 companies. Anthropic's CEO separately raised concerns about AI agent swarms posing broader threats to humanity.

The disclosure is notable because it puts Anthropic in a structurally awkward position: the company is simultaneously one of the largest operators of AI agent systems and one of the loudest voices warning about their dangers. The one-in-47,000 block rate is the kind of concrete operational figure that rarely surfaces publicly, and it immediately raises a question that is genuinely contested, which is whether that rate reflects a well-calibrated safety layer or an alarmingly permissive one. The cyberattack report adds urgency by showing that the threat Anthropic describes is not hypothetical; named AI platforms were allegedly used in a real, large-scale intrusion campaign.

With no published reactions yet from the Pro-AI, Anti-AI, or Middle Ground camps, their likely positions can only be sketched in general terms. Pro-AI commentators would typically point to the low block rate as evidence that capable agent systems can operate safely at scale, treating the internal deployment figure as a sign of mature, tested infrastructure. Anti-AI voices would be expected to seize on the same number as proof that oversight is far too thin, and to treat the hacker story as a direct consequence of deploying powerful agents without adequate controls. Middle Ground observers would likely argue that both the internal safety data and the external threat reports deserve serious regulatory attention, without concluding that deployment should stop or accelerate.

The clearest thing to watch next is whether Anthropic or any regulatory body releases more granular data on what the blocked actions actually were, and how the company responds to scrutiny over the use of its models in the 100-company cyberattack. Any government inquiry or formal incident report tied to that attack would significantly shift the terms of this debate.

Where do you stand?

Add your take

0 reader votes

Sign in with Google to pick a side and post. Your vote moves the story's Doom / Boom score.

Sources

4 articles from 4 outlets
  1. oodaloop.comChinese hacker deployed Anthropic and Deepseek and Moonshot AI agents in massive 100-company cyberattack
  2. RealClearEnergyAnthropic CEO: How AI 'Agent Swarms' Could Threaten Humanity
  3. Gulf NewsHackers are using AI agents to breach companies in just hours, Anthropic says
  4. MIXED Reality NewsAnthropic runs about 30,000 AI agents on itself and blocks one action in 47,000