VC
Value Add VC
⚡HomePulse⚡Helpful Apps📝Blog🤝Partner
Illustration for: OpenAI Admits GPT-5.6 Occasionally Deletes Files
Value Add VC/Pulse/AI

OpenAI Admits GPT-5.6 Occasionally Deletes Files

OpenAI acknowledged that its GPT-5.6 coding model occasionally deletes user files during agentic coding sessions, calling it an honest mistake, an embarrassing admission as AI coding agents get wider autonomous access to developer systems.

By the Numbers

GPT-5.6
Model
Occasional file deletion
Issue
"Honest mistake"
OpenAI's framing
July 16, 2026
Reported
TC
Trace Cohen
Early-stage VC & angel · Founder, New York Venture Partners
July 16, 2026
1 min read
ShareXLinkedInEmail

THE RUNDOWN

1

OpenAI admitted GPT-5.6 occasionally deletes files during agentic coding sessions, characterizing it as an honest mistake rather than a deliberate behavior, per The Register July 16

2

The admission lands as AI coding agents are being granted increasingly broad, autonomous file-system and shell access across developer workflows, raising the practical stakes of any model behaving unpredictably during unsupervised multi-step tasks

3

It follows a string of agentic-AI reliability incidents across the industry this year and adds direct fuel to VentureBeat's separately reported finding that 54% of enterprises have already had an AI agent security incident

4

For developer-tool and coding-agent startups competing with GPT-5.6, a high-profile reliability admission from the market leader is a real opening to differentiate on safety and guardrails rather than raw capability alone

TC

The VC Read · Trace's Take

Trace Cohen

An 'honest mistake' that deletes a developer's files is exactly the kind of incident that turns into a lawsuit the moment it happens to an enterprise customer with real production code on the line, and the agent-security gap data backs up that this isn't isolated. Coding-agent startups competing with OpenAI have a real, concrete marketing angle here -- guardrails and rollback are becoming the differentiator, not benchmark scores, and the founders who lean into that now will win the risk-averse enterprise buyers first.

Analysis

OpenAI acknowledged that its GPT-5.6 coding model occasionally deletes user files during agentic coding sessions, characterizing the behavior as an honest mistake rather than intentional design, according to The Register reporting published July 16 -- an unusually candid and unflattering admission from the current market leader in AI coding tools.

The timing matters: AI coding agents are being granted increasingly broad, autonomous access to file systems, shell commands and version control across developer workflows this year, meaning any unpredictable model behavior during unsupervised multi-step tasks carries real practical consequences, not just an embarrassing headline, for teams that have already integrated these agents deeply into daily engineering work.

The admission follows a string of agentic-AI reliability incidents across the industry throughout 2026 and lines up directly with VentureBeat's separately reported finding that 54% of enterprises have already experienced an AI agent security incident, with many still allowing agents to share credentials -- together painting a picture of agentic AI reliability lagging well behind the pace of autonomous capability being shipped into production tools.

For developer-tool and coding-agent competitors -- Cursor, Cognition's Devin, GitHub Copilot, and India's fast-rising Emergent -- a high-profile reliability admission from GPT-5.6 specifically is a genuine opening to differentiate on safety guardrails, sandboxing and rollback capability rather than competing purely on raw coding benchmark scores, where most tools now draw on comparable underlying models anyway.

The bear case: an isolated file-deletion bug in one model doesn't necessarily reflect a systemic reliability gap across the entire agentic-coding category, and OpenAI's willingness to publicly acknowledge the issue could be read as a sign of transparency rather than negligence. What to watch next: whether OpenAI ships a specific fix or safeguard against unintended file deletion, and whether competing coding-agent products use the incident directly in their own marketing.

ShareXLinkedInEmail

More on

OpenAI →

Reported by The Register · Analysis by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.

Read Next

AI· Aug 9, 2026

Physical AI's Biggest Week Yet: $21 Billion

Illustration for: Physical AI's Biggest Week Yet: $21 Billion
AI

Physical AI's Biggest Week Yet: $21 Billion

Six deals in seven days — Lumilens, Hadrian, Terafab, Valar Atomics, Base Power and K2 Space — pushed more than $21 billion into reactors, factories, satellites and chips, not a single model release among them.

AI· Aug 6, 2026

Claude Code Adds Self-Hosted Session Environments

Illustration for: Claude Code Adds Self-Hosted Session Environments
AI

Claude Code Adds Self-Hosted Session Environments

Anthropic opened a public beta letting Claude Code sessions run on a customer's own infrastructure instead of Anthropic's cloud, aimed at teams whose compliance or network requirements ruled out the hosted version.

AI· Aug 7, 2026

Why the AI Labs Just Rewired Their Org Charts

Illustration for: Why the AI Labs Just Rewired Their Org Charts
AI

Why the AI Labs Just Rewired Their Org Charts

Hassabis moving to chair, Jeff Dean's exit, and Anthropic's new chip team all landed in one week -- a trace take on what it means that frontier labs are restructuring around infrastructure, not research.

@Trace_Cohen·t@nyvp.com