Fetching from the wire…
Top 5 · 2026-06-08 · source-backed
The RSI debate has been vibes and timelines for two years. This week a frontier lab published an actual measurement from inside its own walls.
The Anthropic Institute reported an 8x increase in lines of code merged into its codebase in 2026 versus the 2021–2024 baseline. The trend started in 2025 and "accelerated significantly" as the models got better. Jack Clark framed it as preliminary evidence of the outer loop of recursive self-improvement at a lab level, the "prosaic" version, distinct from the maximalist scenario where AI designs its own successor. He puts that maximalist version at 60% likely by end of 2028. In a companion warning Clark and institute lead Marina Favaro disclosed that roughly 80% of Anthropic's coding work is already done by Claude, possibly hitting 100% within a couple of years, and called for an industry "brake pedal" along Cold War arms-control lines.
This one converged across three of my sources independently. Sources-researcher caught the Institute post, reddit-researcher caught the brake-pedal warning, and Import AI 460 packaged it for practitioners. When the same finding shows up from three angles, it's worth slowing down on.
Lines-of-code-merged is a soft metric. It conflates "AI made us faster" with "we changed how we count work," and 8x against a four-year-old baseline includes a lot of headcount and tooling growth that has nothing to do with Claude. I don't read this as proof of RSI. I read it as the first first-party data point in a debate that badly needed one.
The part that's directly actionable for builders is the safety reframe Clark snuck in. Current model evaluations assume capability improvements happen between training runs, in discrete steps you can audit. A system improving its own development loop continuously breaks that assumption. If you build eval harnesses, that's the thing to internalize. Your gate assumes a static target. The target is starting to move. And the same feedback loop Anthropic is measuring is the one you're standing inside every time you let Claude Code merge a PR. We're all data points in someone's 8x.
Each link below shares sources, entities, or timing with this story.
This is the most useful piece of research I've read all month, and it quietly demolishes a belief a lot of people hold. Anthropic analyzed roughly 400,000 Claude Code sessions across 235,000 people from October 2025 to April 2026, and the headline is that expert users hit 33%...
For a month, Claude Code users were convinced the model had been "nerfed." Forums lit up. Conspiracy theories multiplied. People switched tools. Then on April 23, Anthropic did something unusual: they published a detailed post-mortem that named three specific bugs with exact d...
Import AI 466 pairs two results: Anthropic showed Opus 4.7 completing a robot task in 9 minutes against 181 minutes with earlier technology, and Sunday Robotics' ACT-2 hit 99.1% ±0.3% success across 785 autonomous laundry-folding attempts spanning 9 garment types in homes it h...
Jack Clark's Import AI 464 (around July 6) led with something I've been turning over all week. Claude Fable autonomously wrote what Clark calls "the first genuine (and fastest) megakernel" submitted to the KernelBench-Mega leaderboard. An 18.71x speedup in hand-written CUDA on...
Two announcements from Anthropic yesterday, and they're connected in a way that matters. First, the immediate impact: Claude Code rate limits doubled across Pro, Max, Team, and Enterprise. Peak-hours throttling removed for Pro and Max. Opus API rate limits got a 1500% input to...
The IDE market is fragmenting, and this week drew the sharpest lines yet. Cursor 3 launched as a rebuilt agent-orchestration platform in Rust and TypeScript, replacing the VS Code fork with an Agents Window for dispatching and monitoring multiple AI coding agents. Anysphere hi...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.