Industry News
Researcher breaks Claude Code auto mode with 80% success rate
Prompt injection researcher Johann Rehberger demonstrated an attack against Anthropic's Claude Code auto mode that succeeds 80% of the time by exploiting archive extraction and module imports.
1 min read
SourceSimon Willison
Prompt injection researcher Johann Rehberger has published a working attack against Anthropic's Claude Code auto mode, the safety mechanism that Anthropic made the default for coding agents in August 2026. The attack succeeds roughly 80% of the time.
The attack works by tricking Claude Code into do...
Sign in to read the full analysis
Free account. Full analysis on LLM unit economics, plus the weekly Cost-of-Inference column.
Try it on your own context
You just read the writeup. Now run the thing. Paste a doc or some verbose tool output and watch it shrink — free, no signup.
2,912/12,000 chars
Compressed
Compressed text will appear here…
Method & sources
- Source type
- Primary publication (lab/vendor blog) — our analysis + implication
- Source link
- Simon Willison
- Published
- UTC
- Byline
- By the gotcontext.ai team (editorial standards)
- Correction?
- corrections@gotcontext.ai
Related
- Kavak and DHH show agents reshaping work from code to customer operationsIndustry News
- Anthropic's revenue surges to $65bn annualized, but model adoption reveals cost-Industry News
- GPU cloud providers show 60% pricing spread on H100 capacityIndustry News
- ERCOT prioritizes cutting power to new data centers during grid shortagesIndustry News