Fetching from the wire…
Public story · 2026-08-24 · high
Breunig splits his own work between Fable for design and GLM 5.2, priced at about a ninth as much, for execution.
Why now: Breunig published the post August 23, and Simon Willison picked it up within a day.
Drew Breunig argued that Fable's release ends a free ride AI engineers have had for years, in a post published August 23. The stakes are concrete for anyone routing work across models. GLM 5.2 launched the same month at about a ninth of Fable's price and a fifth of Opus 5's, per Breunig.
Sutter's 2005 essay, 'The Free Lunch Is Over,' supplies the comparison. CPU clock speeds doubled every 18 months for decades, so optimizing code was a bad bet. Faster hardware would fix a slow program within a year or two.
When clock scaling stalled, optimization stopped being optional.
The model-release cycle did the same job for years, he says. Tuning a harness or a context strategy didn't pay off, because a cheaper, better model would show up in a few months. It would cover for whatever you'd built.
Fable changed that math. It's strong and priced high enough that teams now have to decide what work goes where. Waiting for the next release to close the gap doesn't work anymore.
His workflow shows how it works. Fable shapes the design, then a written brief goes to GLM 5.2 to execute it.
Its access controls and data-retention rules will lock this multi-model split in, Breunig expects, since companies will start caring where their traces go.
Ramp's numbers already show developers running a portfolio of models instead of one default. Breunig's post argues the reason is that price-per-dollar model gains slowed, so the payoff shifted to everything built around the model.
Breunig doesn't know if the split holds past the next release cycle, and says as much in the post.
Each link below shares sources, entities, or timing with this story.
$3,054 against $38,370. Same benchmark, better score. Praxist (arXiv 2608.25955, submitted August 26) replaces per-attempt agent memory with a typed evidence graph of findings, plus lane-structured frontiers and agendas, so later attempts inherit validated mechanisms rather th...
Simon Willison pulled the numbers out of an FT report sourced to "people with knowledge of the matter": Anthropic's annualized revenue reached $65bn in July, up from $47bn in May. Six thousand customers spend $100,000 or more a year. The company told investors it expects a pro...
Terminal-Bench 2.1 results (entries dated June 17) put Codex CLI on GPT-5.5 first at 83.4%, Claude Code on Fable 5 second at 83.1%, and Claude Code on Opus 4.8 at 78.9%. The asterisk matters more than the ranking: Fable 5 and Mythos 5 have been export-suspended since June 12,...
Spotify's Portal team published Xirp on August 10: a vendor-neutral agentic development environment that manages concurrent sessions across Claude Code, Gemini CLI, and Codex, each session isolated in its own git worktree so dozens of agents can work the same codebase without...
Everyone spent yesterday arguing about benchmark numbers. Tencent quietly published data suggesting the numbers belong to your infrastructure, not the model. The WorkBuddy Bench leaderboard reports every model under two different agent harnesses — CodeBuddy Code and Claude Cod...
This is the other half of the Fable 5 story, so read them together. While the best coding model in the world is uncallable, an open-weight one quietly posted frontier-adjacent numbers. Per Tom's Hardware, independent benchmarks for the MIT-licensed GLM-5.2 (744B params, 40B ac...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.