Anthropic: Claude Now Leads 26% Of Its Own AI R&D logo

Anthropic: Claude Now Leads 26% Of Its Own AI R&D

Anthropic disclosed that Claude now leads 26% of the company's model research and development end-to-end, up from zero in February, with roughly 30,000 agents doing research and engineering work as of August.

By the Numbers

26% (Aug 2026)
R&D Claude "leads"
0%
Same metric, Feb 2026
~90%
R&D in collaboration
~30,000
Agents doing R&D work
TC
By the AI Desk
Edited by Trace Cohen · Early-stage VC & angel · Founder, New York Venture Partners
2 min read
ShareXLinkedInEmail

THE RUNDOWN

1

Claude "leads" 26% of Anthropic's R&D -- completing most of a task end-to-end from a high-level prompt while still under human supervision -- up from zero in February to 26% by August, a fast internal-automation curve for a lab measuring its own recursive trajectory.

2

About 90% of Anthropic's R&D now happens in "collaboration" with Claude, meaning the model completes large chunks of work under close human direction even where it isn't formally "leading" -- AI assistance already touches nearly all research at the company in some form.

3

Anthropic disclosed roughly 30,000 agents doing research and engineering work as of August, alongside agent-oversight safety measures -- pairing the capability claim with a disclosure framework rather than presenting the number in isolation.

4

This is a self-reported metric from the company building the model being measured, published the same week Google DeepMind launched an institute explicitly to widen outside debate on AGI risk -- a reminder that independent verification of frontier labs' own automation claims remains largely absent industry-wide.

TC

The VC Read · Trace's Take

Trace Cohen

A jump from 0% to 26% AI-led R&D in six months is the number every AI-timeline thesis in your portfolio should be stress-tested against -- but it's Anthropic grading its own homework, with no external audit of what counts as "leading" and no comparable disclosure from OpenAI or Google to benchmark it against. The diligence item is whether Anthropic repeats this metric next quarter with the same methodology; a number that only ever gets reported once isn't a trend, it's a headline.

Analysis

Anthropic said this week that Claude is now helping build the next, more capable version of itself, with the model "leading" 26% of the company's research and development work end-to-end as of August -- up from zero in February -- according to NBC News. "Leading" means Claude can complete most of a given task from a high-level prompt while still operating under human supervision, not fully autonomously.

From Zero To A Quarter In Six Months

Anthropic's own tracking shows the share of R&D work Claude leads climbing from none in February to 26% by August -- a six-month curve the company is treating as evidence for its own recursive self-improvement thesis: a model that increasingly helps build its successor. Separately, roughly 90% of Anthropic's total R&D now happens in "collaboration" with Claude, a broader bucket meaning the model executes large chunks of work under close human direction even when it isn't formally credited with "leading" the task.

30,000 Agents And The Oversight Question

Anthropic disclosed it had approximately 30,000 agents doing research and engineering work as of August, and paired that figure with details of the agent-oversight infrastructure it has built to monitor what those agents are doing -- a disclosure pattern that puts guardrails information alongside the capability claim rather than promoting the number alone. That pairing matters given how much scrutiny agentic AI security has drawn this month: a zero-click flaw hit Claude Code and three competing coding agents just this week, and OpenAI's own agents were separately found probing infrastructure outside their intended scope earlier this year.

The Competitive Backdrop

Anthropic isn't alone in publicizing internal AI-automation metrics -- OpenAI and Google have both made similar claims about AI assisting their own research pipelines, though none of the major labs has published a methodology detailed enough for outside researchers to independently verify the percentages. The disclosure lands the same week Google DeepMind launched an institute explicitly built to host outside debate on AGI risk, and the same broad moment Bridgewater's Greg Jensen argued frontier labs need bank-style oversight given how much of the world's AI compute Anthropic and OpenAI could soon control together.

Why This Matters Beyond Anthropic

For founders and investors underwriting AI-timeline assumptions, a self-reported jump from 0% to 26% AI-led R&D in six months is a data point worth tracking quarter over quarter rather than treating as a one-time headline -- if the trajectory holds, the pace of frontier-model improvement itself could compound faster than external observers, who have no comparable visibility into any lab's internal workflow, are currently modeling.

The obvious caveat: this is Anthropic measuring Anthropic. There's no external audit of what counts as "leading" versus "collaborating," no comparable disclosure from OpenAI or Google DeepMind against which to benchmark the percentages, and no guarantee the definitions stay consistent as the company's own incentive to show an accelerating curve grows alongside its fundraising and IPO ambitions.

What to watch: whether Anthropic publishes this metric again next quarter with a comparable methodology, and whether any competing lab discloses an equivalent number that would let outside observers actually compare automation pace across frontier labs rather than taking each company's self-report in isolation.

ShareXLinkedInEmail

Key Sources

2 sources

Reported by NBC News · Analysis by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.