Anthropic says 26% of its own R&D is now done by Claude, running 30,000 internal agents at a time
Anthropic's new transparency index shows roughly 30,000 Claude agents working inside the company at any moment, doing 26% of its AI research and engineering work, with a real-time monitor blocking about 1 in 47,000 of over a billion agent decisions logged in August — the clearest public data yet on how far AI-assisted development has scaled inside a frontier lab.
28 September 2026
Anthropic just published the first numbers from what it’s calling a measurement of “the pace of AI development inside frontier labs,” and the headline figures are striking: at any given moment, around 30,000 Claude agents are running research and engineering tasks on Anthropic’s own internal platform, and the company estimates that 26% of its AI R&D work is now done by Claude rather than a human engineer. Roughly 6% of the company’s AI R&D compute is spent specifically on safety monitoring for that agent fleet.
The oversight numbers matter more than the scale
The eye-catching part isn’t the headcount of agents — it’s what happens around them. Anthropic runs two layers of monitoring: a real-time check that screens every agent action before it executes, and a slower post-hoc review for anything more subtle. Across more than a billion agent decisions logged in August, the real-time monitor blocked about 0.002% of them — roughly 1 in 47,000 — mostly actions with fast, hard-to-reverse consequences, like copying model weights somewhere they shouldn’t go. The offline layer flags around 100,000 records a week for automated classification, and about 50 of the highest-priority cases get escalated to a human reviewer.
Why this is a search signal, not just a curiosity
This is the first time a frontier AI lab has published concrete numbers on its own internal agent-to-human R&D ratio and the guardrails wrapped around it, and it’s already reframing a question we hear constantly from clients: “if AI is writing this much of the code, who’s actually watching it?” Anthropic’s answer — every action gated before execution, with tiered escalation for anything ambiguous — is effectively the production version of the review discipline any serious AI-assisted development engagement should already have, just running at a scale most teams will never need.
So what
The lesson for anyone commissioning software isn’t “let the agents run unsupervised because Anthropic does” — it’s the opposite. Anthropic’s own numbers show that scaling agentic development safely means building explicit, layered oversight into the process, not assuming a capable model makes review optional. That’s exactly the posture we take on AI-assisted development projects: AI agents move fast on the parts of a build that suit them, and a human still owns every decision that’s hard to reverse. If you’re weighing how much of your next build should run through AI tooling versus a person, get in touch and we’ll walk you through where that line sits for your project.