VC
Value Add VC
⚡HomePulse⚡Helpful Apps📝Blog🤝Partner
Illustration for: DeepSeek Ships V4 Pro, Its Sharpest Agent Model Yet
Value Add VC/Pulse/AIDEEP DIVE

DeepSeek Ships V4 Pro, Its Sharpest Agent Model Yet

DeepSeek took V4 Pro 0813 to general availability with sharply improved agentic benchmarks, but the vendor-reported gains haven't been independently replicated and a price hike lands within days.

By the Numbers

87.9 (was 72.1)
Terminal Bench 2.1
62.7 (was 12.8)
DeepSWE score
1M tokens
Context window
384K tokens
Max output
Aug 16, 16:00 UTC
Price hike effective
TC
By the AI Desk
Edited by Trace Cohen · Early-stage VC & angel · Founder, New York Venture Partners
August 13, 2026
3 min read
ShareXLinkedInEmail

THE RUNDOWN

1

DeepSeek moved DeepSeek-V4-Pro-0813 to general availability, updating its API pricing page on August 12 -- the model targets agentic tasks where AI systems use tools, run code, and finish multi-step workflows without a human in the loop, per [Tech Times](https://www.techtimes.com/articles/324241/20260813/deepseek-v4-pro-0813-goes-ga-benchmark-claims-await-independent-proof.htm)

2

DeepSeek's own benchmarks show large jumps versus the preview build: Terminal Bench 2.1 rose from 72.1 to 87.9, DeepSWE from 12.8 to 62.7, and CyberGym from 52.7 to 83.3 -- gains as large as 49.9 percentage points that no third-party evaluator has yet reproduced

3

The model handles up to a 1 million-token context window and can output as much as 384,000 tokens, with a switch between thinking and non-thinking modes -- and DeepSeek has told developers a price increase for the V4 family takes effect at 16:00 UTC on August 16, just days after this release

4

The launch follows [Pulse's coverage](/pulse/deepseek-50-billion-yuan-funding-price-hikes-2026) of DeepSeek resuming a funding round targeting roughly 500 billion yuan (about $70 billion) in pre-money valuation, expected to close in late August -- a model release timed just ahead of both a price hike and a funding close

TC

The VC Read · Trace's Take

Trace Cohen

The diligence item isn't the benchmark table, it's that every number on it is self-reported -- I'd wait for Artificial Analysis or a comparable third party to run Terminal Bench and DeepSWE independently before treating a 50-point jump as real. The timing tells its own story: shipping a stronger model four days before a price hike is a usage-lock tactic, not just a capability release. Watch the ~$70B pre-money funding round DeepSeek is targeting to close this month -- a strong independent benchmark confirmation right before that close would be the more interesting signal than the release itself.

AI Landscape →

Analysis

The Release

DeepSeek pushed DeepSeek-V4-Pro-0813 to general availability, updating its official API pricing page on August 12 to reflect the new build, according to Tech Times and corroborated by Yahoo. The GA release focuses on agentic capability -- tasks where a model uses tools, writes and executes code, and completes multi-step workflows without a human stepping in at each turn, rather than chasing a higher score on general chat benchmarks.

The Numbers DeepSeek Is Showing

DeepSeek's self-reported benchmarks show some of the largest single-release jumps of any model this year:

  • Terminal Bench 2.1 -- 87.9, up from 72.1 in the preview build
  • DeepSWE -- 62.7, up from 12.8
  • CyberGym -- 83.3, up from 52.7
  • NL2Repo -- 61.5

Those are gains of up to 49.9 percentage points versus the preview. The model supports a context window up to 1 million tokens, can generate outputs as long as 384,000 tokens, and lets developers toggle between a thinking mode and a faster non-thinking mode depending on the task.

Company Background

DeepSeek was founded in 2023, spun out of the Chinese quantitative hedge fund High-Flyer, and built its early reputation on delivering frontier-adjacent capability at a fraction of the compute cost US labs assumed was necessary. Pulse covered DeepSeek's V4-Flash costing roughly 3 cents per benchmark test to run -- more than 100 times cheaper than Anthropic's Claude Fable 5, albeit at meaningfully lower intelligence scores.

The funding trajectory has moved just as fast as the model releases:

  • June 2026 -- closed a $7.4 billion round backed by Tencent, CATL, JD.com, NetEase, and IDG Capital at a valuation north of $50 billion
  • July 2026 -- moved into fresh talks for a raise at roughly $71 billion, Pulse reported
  • This month -- Pulse most recently reported DeepSeek resumed a round targeting about 50 billion yuan at a pre-money valuation near 500 billion yuan, expected to close late August -- the same window this release lands in

The Competitive Field

V4 Pro's agentic focus puts it in direct competition with OpenAI's GPT-5.6 family, Anthropic's Claude Code line, and Google's Gemini 3.6 Flash on tool-use and coding-agent benchmarks -- the same category Cognition's Devin and Cursor-maker Anysphere (now part of SpaceX) compete in as products rather than base models. Domestically, DeepSeek is racing Alibaba's Qwen and Moonshot AI's Kimi, both of which Pulse has covered as also preparing IPO paths this year.

The Counterweight

Every one of the headline benchmark gains is vendor-reported, and that's a real limitation worth noting before treating the numbers as settled. No independent evaluator -- not Artificial Analysis, not a third-party lab -- has yet replicated DeepSeek's Terminal Bench or DeepSWE numbers for this specific build, and a 49.9-percentage-point jump on DeepSWE in particular is large enough that it deserves outside verification before it's treated as fact rather than marketing. Releasing a stronger model just four days before a price increase also isn't neutral timing: it gives developers a reason to lock in usage at current rates before the August 16 hike, which is a demand-pull tactic as much as a pure capability story.

Watch whether independent benchmarks confirm the jump once V4 Pro sees wider use, and whether the August 16 price increase changes DeepSeek's cost advantage over Western frontier models now that the gap has narrowed.

ShareXLinkedInEmail

More on

DeepSeek →

Reported by Tech Times · First reported by Yahoo Tech · Analysis by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.

Read Next

AI· Aug 13, 2026

Anthropic in Talks to Buy AI Video Startup Decart for $6B

Illustration for: Anthropic in Talks to Buy AI Video Startup Decart for $6B
AI~$6B acquisition talks

Anthropic in Talks to Buy AI Video Startup Decart for $6B

Anthropic is negotiating to buy Israeli startup Decart, which builds real-time video-generation models and GPU-efficiency software, for roughly $6 billion in what would be Anthropic's largest acquisition ever.

AI· Aug 12, 2026

Google DeepMind's Talent Exodus Reveals a Deeper Identity Crisis

Illustration for: Google DeepMind's Talent Exodus Reveals a Deeper Identity Crisis
AI

Google DeepMind's Talent Exodus Reveals a Deeper Identity Crisis

Days after Demis Hassabis stepped down as DeepMind CEO, Fortune reports stalled models, missed deadlines and staff burnout drove the exodus, with engineers telling the outlet DeepMind is losing its separation and identity within Alphabet.

AI· Aug 13, 2026

Tech CEOs Are Writing Manifestos Nobody Asked For

Illustration for: Tech CEOs Are Writing Manifestos Nobody Asked For
AI

Tech CEOs Are Writing Manifestos Nobody Asked For

Fortune examines a wave of lengthy personal essays from tech CEOs -- some running past 6,500 words -- on AI's future, raising the question of who reads them and whether they function as genuine philosophy or investor signaling.

@Trace_Cohen·t@nyvp.com