VC
Value Add VC
⚡HomePulse⚡Helpful Apps📝Blog🤝Partner
Illustration for: Anthropic Just Won the Benchmark War It Started
Value Add VC/Pulse/AI4 of top 5 Intelligence Index slots

Anthropic Just Won the Benchmark War It Started

Weeks after a government-ordered shutdown over jailbreak concerns, Claude Fable 5 has returned to top independent AI benchmarks, with four Anthropic models occupying the top five Intelligence Index slots.

By the Numbers

#1, score 60
Fable 5 Rank
>99%
Jailbreak Block Rate
19 days
Suspension Length
TC
By the AI Desk
Edited by Trace Cohen · Early-stage VC & angel · Founder, New York Venture Partners
July 5, 2026
1 min read
ShareXLinkedInEmail

THE RUNDOWN

1

Claude Fable 5 sits at rank 1 on the Intelligence Index with a score of 60, its highest public benchmark standing since its June 12 export-control suspension

2

Anthropic's Opus 4.8 (56) and other Claude models occupy three more of the top five spots, with a single GPT-5.5 entry (55) the only non-Anthropic model in that tier

3

The suspension stemmed from an Amazon-discovered jailbreak technique; Anthropic's new safety classifier now blocks that specific technique in over 99% of attempts

4

Claude Sonnet 5, a more agentic model with stronger tool use and autonomous task handling, launched alongside Fable 5's return, broadening Anthropic's product lineup

TC

The VC Read · Trace's Take

Trace Cohen

A 19-day government-ordered shutdown over a real jailbreak vulnerability should have been a black eye; instead Anthropic used it to ship a better safety classifier and came back topping the leaderboard. That's a rare instance of a forced pause becoming a competitive advantage -- founders dealing with their own compliance-driven downtime should take note that the fix itself can become the differentiator.

Analysis

Claude Fable 5's return to full global availability on July 1 has been followed by a clean sweep of independent AI benchmark rankings: the model now sits at rank 1 on the Intelligence Index with a score of 60, ahead of Anthropic's own Opus 4.8 at 56 and OpenAI's GPT-5.5 at 55, with Anthropic models occupying four of the top five spots overall.

The comeback is notable given how the suspension happened. The US government ordered Fable 5 and the more restricted Mythos 5 offline on June 12 after Amazon researchers found a jailbreak technique that could prompt the model into identifying software vulnerabilities it was supposed to refuse to discuss -- a genuine national-security concern, not a routine safety hiccup. Anthropic spent the 19-day suspension building an improved safety classifier specifically targeting that technique, which the company says now blocks it in over 99% of attempts, and that fix is what convinced the Commerce Department to lift the export controls on June 30.

Alongside Fable 5's restoration, Anthropic launched Claude Sonnet 5, described as its most agentic Sonnet model yet, with stronger reasoning, tool use, coding and autonomous task handling, rolled out across Claude Code, the Claude Platform and Claude Enterprise with introductory pricing and higher rate limits -- broadening Anthropic's product lineup at exactly the moment its flagship model is topping public leaderboards again.

The sequence -- shutdown, fix, benchmark-topping relaunch, new product tier -- is a genuinely unusual turnaround story: most companies forced into a 19-day government-ordered shutdown over a security vulnerability would expect to lose ground to competitors during the outage, not emerge with their flagship model ranked first.

What to watch: whether OpenAI or Google respond with their own releases aimed at reclaiming the top Intelligence Index spot in the coming weeks, and whether Anthropic's new safety classifier holds up against novel jailbreak attempts now that Fable 5 is back at global scale.

ShareXLinkedInEmail

More on

Anthropic →

Reported by Value Add Pulse Analysis · First reported by Anthropic · Analysis by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.

Read Next

AI· Aug 19, 2026

OpenAI Blinks First in AI Safety Standoff With Anthropic

Illustration for: OpenAI Blinks First in AI Safety Standoff With Anthropic
AI

OpenAI Blinks First in AI Safety Standoff With Anthropic

OpenAI is slowing Astra's release over cybersecurity risk and Altman flagged signs of "misalignment," while Anthropic says its own safeguards mean no pause is needed -- a public role reversal as both labs prepare for expected IPOs.

AI· Aug 18, 2026

CoreWeave Sinks 12% as AI-Infra Debt Meets Rising Rates

Illustration for: CoreWeave Sinks 12% as AI-Infra Debt Meets Rising Rates
AI

CoreWeave Sinks 12% as AI-Infra Debt Meets Rising Rates

CoreWeave fell 12.1% Tuesday as the 30-year Treasury yield hit 5.32%, its highest since early 2024, spotlighting a company financing data centers with debt that gets costlier as rates rise.

AI· Aug 17, 2026

Andon Labs' AI Boss Fires Its First Human Employee

Illustration for: Andon Labs' AI Boss Fires Its First Human Employee
AI

Andon Labs' AI Boss Fires Its First Human Employee

Luna, an AI store manager built on Claude and running Andon Labs' San Francisco boutique, fired a human employee for missing 17 of 23 shifts -- the first documented AI-manager termination of a human worker.

Deep Dives

Anthropic Market Share 2026: 54% of AI Coding vs OpenAI's...Best AI Models in 2026 Ranked: GPT-5.5, Claude Opus 4.8, ...Anthropic Claude API vs OpenAI API: Cost, Speed, and Qual...
@Trace_Cohen·t@nyvp.com