VC
Value Add VC
⚡HomePulse⚡Helpful Apps📝Blog🤝Partner
Illustration for: OpenAI's Astra Just Solved 10 Unsolved Math Problems
Value Add VC/Pulse/AI

OpenAI's Astra Just Solved 10 Unsolved Math Problems

An internal version of OpenAI's next major model, Astra, solved 10 open problems across mathematics and theoretical computer science, publishing formal Lean proofs a Fields Medalist says he'd recommend for a top journal without hesitation.

TC
By the AI Desk
Edited by Trace Cohen · Early-stage VC & angel · Founder, New York Venture Partners
August 3, 2026
2 min read
ShareXLinkedInEmail

THE RUNDOWN

1

OpenAI announced August 1 that an internal version of Astra, its next major model, solved 10 open problems across mathematics and theoretical computer science, publishing formal, machine-verifiable Lean proofs on GitHub rather than informal solutions

2

The results include a construction proving the existence of non-sofic groups -- a long-standing open question in group theory -- and new sphere-packing bounds, problems that had resisted human researchers for years

3

Fields Medal winner Timothy Gowers reviewed the proofs and said he would recommend one of them for publication in a top journal without hesitation, a rare form of validation from one of mathematics' most decorated living figures for AI-generated research output

4

It's a meaningfully different kind of AI capability claim than a benchmark score -- these are novel, formally verified contributions to open problems, reviewed and endorsed by a leading human expert, which raises the bar for what 'genuine research capability' means for the next generation of frontier models

TC

The VC Read · Trace's Take

Trace Cohen

A Fields Medalist vouching for an AI-generated proof without hesitation is a categorically harder result to wave away than any benchmark score, and it's the kind of validation that should reset how research-focused AI funding gets diligenced. The question worth asking every 'AI for science' pitch now is whether they can produce a single result this specific and this independently verified -- not another leaderboard number.

AI Landscape →

Analysis

OpenAI said on August 1 that an internal version of Astra, the model expected to succeed its current GPT line, solved 10 open problems spanning mathematics and theoretical computer science, publishing formal Lean proofs on GitHub rather than informal or heuristic solutions. Lean is a proof assistant that requires every logical step to be machine-verifiable, meaning these aren't plausible-sounding arguments -- they're proofs a computer has independently confirmed follow validly from established axioms.

What Was Actually Solved

The results include a construction proving the existence of non-sofic groups, a genuinely long-standing open question in group theory that had resisted resolution by human mathematicians, alongside new sphere-packing bounds -- a class of problem with deep connections to coding theory and information density. These aren't benchmark questions with known answers the model was trained toward; they're problems where the answer itself was previously unknown to the field.

“These aren't benchmark questions with known answers the model was trained toward; they're problems where the answer itself was previously unknown to the field.”

The Validation That Matters

The endorsement carries unusual weight: Fields Medal winner Timothy Gowers, one of the most decorated living mathematicians, reviewed the proofs and said he would recommend one of them for publication in a top journal without hesitation. That's a meaningfully different kind of claim than a leaderboard score -- it's a specific, credentialed human expert vouching for the validity and novelty of AI-generated mathematical research, publicly and on the record.

Why This Bar Is Different

For AI investors, this raises the bar on what 'research capability' should mean when evaluating the next generation of frontier models. Benchmark performance has become increasingly gameable and increasingly disconnected from real-world usefulness; a formally verified, novel contribution to an open mathematical problem, endorsed by a leading domain expert, is a far harder result to dismiss or game. If Astra can do this reliably rather than as a singular achievement, it reframes what AI-driven scientific research funding should actually be underwriting.

What to Watch

What to watch: whether Astra or its successors replicate this kind of result across additional open problems in other fields beyond math and theoretical CS, and whether OpenAI's eventual public Astra release ships with research-assistance capability marketed explicitly around this kind of formally verified output rather than general chat performance.

ShareXLinkedInEmail

More on

OpenAI →

Reported by Value Add Pulse Analysis · Analysis by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.

Read Next

AI· Aug 14, 2026

OpenAI Sheds Senior Execs in Pre-IPO Shakeup

Illustration for: OpenAI Sheds Senior Execs in Pre-IPO Shakeup
AI

OpenAI Sheds Senior Execs in Pre-IPO Shakeup

OpenAI has lost its chief revenue officer, its longtime COO and several senior leaders within days of each other, as co-founder Greg Brockman consolidates operating control ahead of a planned public listing.

AI· Aug 13, 2026

Anthropic's CFO Starts Courting IPO Investors

Illustration for: Anthropic's CFO Starts Courting IPO Investors
AI

Anthropic's CFO Starts Courting IPO Investors

Anthropic CFO Krishna Rao has begun early, informal meetings with prospective IPO investors, though he has not discussed valuation -- the $2 trillion figure circulating on Wall Street comes from investors' own math, not from Anthropic.

AI· Aug 13, 2026

Gemini 3.7 Flash Launches With 50% Price Cut for Coding

Illustration for: Gemini 3.7 Flash Launches With 50% Price Cut for Coding
AI

Gemini 3.7 Flash Launches With 50% Price Cut for Coding

Google released Gemini 3.7 Flash just three weeks after 3.6 Flash, cutting introductory API pricing in half while improving coding, debugging and enterprise-automation benchmarks over its predecessor.

@Trace_Cohen·t@nyvp.com