Fetching from the wire…
Vibe Coding2026-08-29 · source-backed
A builder who kept burning weekly Fable limits documented testing Fable with Opus, Grok, K3, Sol high/xhigh, Sol-as-orchestrator via Codex and Opus-as-orchestrator, finding most combinations land near the worker model's quality rather than the orchestrator's. Cranking Sol to max thinking as the worker was the only pairing holding sustained Fable-level quality, moving consumption from roughly 50% of the weekly Fable budget in a few hours to about 10% Fable plus 10-15% Codex. The caveat is in the post: it requires two expensive subscriptions. "Output quality tracks the worker, not the orchestrator" is a finding I'd want replicated, because it contradicts how most people describe orchestration. (r/ClaudeAI)
Each link below shares sources, entities, or timing with this story.
Terminal-Bench 2.1 results (entries dated June 17) put Codex CLI on GPT-5.5 first at 83.4%, Claude Code on Fable 5 second at 83.1%, and Claude Code on Opus 4.8 at 78.9%. The asterisk matters more than the ranking: Fable 5 and Mythos 5 have been export-suspended since June 12,...
Fable 5.1 came out this week and two people independently measured what it costs. They disagree by a factor of about four, and both are right. A MineBench run of 15 identical Minecraft builds put Fable 5.1 at $147.55 total against Fable 5's $54.93. Average inference time went...
OpenAI launched the GPT-5.6 family on July 14: Sol (flagship), Terra (cost-optimized), and Luna (fast tier), live across ChatGPT, Codex, and the API the same day after a US-government-requested delay for security review. The numbers are loud. Sol scored 53.6 on Agents' Last Ex...
Everyone kept score wrong. When OpenAI shipped GPT-5.6 (the Sol flagship plus Terra and Luna) to GA on July 9, then xAI put out Grok 4.5, Meta dropped Muse Spark 1.1, and Cognition shipped SWE-1.7, the reflex was to ask who won the benchmark. Wrong question. On the Artificial...
SpaceXAI released Grok 4.5 on July 8, and for once the vendor hype and the third-party numbers point roughly the same direction. Musk called it "roughly comparable to Opus 4.7, but much faster." Priced at $2 per million input tokens and $6 per million output, that's over 60% b...
Per-token prices went down at both labs. Subscriptions are draining faster at both labs. Those aren't in tension once you look at token counts. On the OpenAI side, r/OpenAI collected reports from Linux.do and NodeSeek alleging Astra consumes more Plus quota than its published...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.