VC
Value Add VC
⚡HomePulse⚡Helpful Apps📝Blog
Illustration for: Microsoft Launches In-House Rival To Anthropic's Mythos
Value Add VC/Pulse/AI96% CyberGym

Microsoft Launches In-House Rival To Anthropic's Mythos

Microsoft unveiled MAI-Cyber-1-Flash, an in-house cybersecurity model that scores 96% on the CyberGym benchmark, beating Anthropic's Mythos, Gemini and GPT while cutting costs roughly in half.

96%
CyberGym score
~50%
Cost reduction
Nov 3, 2026
Preview date
TC
Trace Cohen
Early-stage VC & angel · Founder, New York Venture Partners
July 27, 2026
1 min read
ShareXLinkedInEmail

THE RUNDOWN

1

Microsoft's MAI division built MAI-Cyber-1-Flash, a compact security model embedded inside MDASH, the company's multi-agent harness for finding and fixing software vulnerabilities

2

The model scores 96% on CyberGym, a benchmark measuring how well AI systems reason over large codebases to find real vulnerabilities, beating frontier competitors including Anthropic's Mythos, Gemini and GPT while cutting inference costs roughly in half versus Microsoft's own prior production setup

3

Microsoft is separately preparing Project Perception, a broader AI security platform that routes tasks across Microsoft, OpenAI and Anthropic models -- reserving expensive frontier calls for high-value steps and assigning cheaper distilled models to high-volume scanning

4

The moves are a direct cost-and-access argument against Anthropic's Claude Mythos Preview, positioning Microsoft as both a customer of and competitor to the frontier labs whose models it also resells through Azure

TC

The VC Read · Trace's Take

Trace Cohen

Microsoft building an in-house model explicitly to undercut the same labs it resells through Azure is the clearest evidence yet that hyperscalers view frontier-lab API margins as fair game to compete away, not just a revenue line to protect. If a distilled, purpose-built model really beats general frontier models on a narrow task at half the cost, that's the template every enterprise software category eventually follows -- and it's bad news for labs pricing on general-purpose capability alone.

Frontier AI Dashboard →

Analysis

Microsoft unveiled its first in-house cybersecurity AI model Monday, MAI-Cyber-1-Flash, built by its Microsoft AI division and embedded inside MDASH, the company's multi-agent system for finding and fixing software vulnerabilities -- a direct challenge to Anthropic's Claude Mythos Preview, which has emerged as a leading model for AI-driven security research.

Microsoft's benchmark claims are aggressive: the company says MAI-Cyber-1-Flash scores 96% on CyberGym, a benchmark that measures how well AI systems reason over large codebases to find genuine, exploitable vulnerabilities, beating frontier competitors including Mythos, Gemini and GPT -- while cutting inference costs roughly in half compared to Microsoft's own previous production configuration.

The compact model is one piece of a broader security push. Microsoft is separately preparing to launch Project Perception, an AI-powered platform that scans enterprise codebases for exploitable vulnerabilities, built around a deliberate cost-and-access argument against Anthropic's positioning. Project Perception routes security analysis tasks across models from Microsoft, OpenAI and Anthropic using a model-selection layer that reserves expensive frontier calls for the steps that actually require them, assigning cheaper, distilled models to high-volume scanning passes. The tools enter preview November 3.

The competitive framing is notable because Microsoft occupies an unusual dual role here: it resells Anthropic's and OpenAI's models through Azure while simultaneously building in-house models designed to outperform and undercut them on price for specific tasks. That's a similar structure to Amazon's relationship with the AI labs it hosts, but Microsoft's security-specific push is more directly adversarial in its framing than most hyperscaler in-house model efforts to date.

What to watch: whether independent benchmarks confirm Microsoft's CyberGym claims once the model is broadly available, how Anthropic responds given Mythos was explicitly positioned as a security-research leader, and whether enterprise security teams actually shift spend toward Microsoft's cheaper option once Project Perception exits preview in November.

ShareXLinkedInEmail
More onAnthropic →Google →Microsoft →

Analysis and editorial commentary by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.

Read Next

AI· Jul 28, 2026

Anthropic's Claude Code Reigns Despite Codex Rise

Illustration for: Anthropic's Claude Code Reigns Despite Codex Rise
AI~42% share

Anthropic's Claude Code Reigns Despite Codex Rise

Claude Code holds roughly 42% of enterprise AI-coding market share, more than double OpenAI's, even as Codex's open-source Apache-2.0 release and OpenCode both gain developer interest.

AI· Jul 28, 2026

Moonshot Seeks More Blackwell Chips For Next Model

Illustration for: Moonshot Seeks More Blackwell Chips For Next Model
AI

Moonshot Seeks More Blackwell Chips For Next Model

Chinese lab Moonshot AI is seeking additional Nvidia Blackwell chips to train Kimi K4, its next model, days after the White House accused it of illegally accessing banned GB300 chips through Thailand-based infrastructure.

AI· Jul 27, 2026

ASML Slides 8% On Report China Mass-Produces DUV Tools

Illustration for: ASML Slides 8% On Report China Mass-Produces DUV Tools
AI-8% ASML

ASML Slides 8% On Report China Mass-Produces DUV Tools

ASML shares fell more than 8% after The Information reported a Shanghai-based company has begun mass-producing immersion DUV lithography tools, the exact chipmaking equipment Dutch and US export controls were designed to keep out of China.

@Trace_Cohen·t@nyvp.com