Fetching from the wire…
Models2026-08-14 · source-backed
Announced August 13, up to 14x Standard processing, 11x faster than Claude Fable 5, 5x faster than Opus 4.8 on Fast mode, running on the Cerebras Wafer-Scale Engine with 44 GB of on-chip SRAM per wafer. Limited API preview for select customers with capacity-gated expansion. The concrete number from Cerebras: all 2,500 Humanity's Last Exam questions in 11 hours 11 minutes versus 78+ hours for Claude Fable 5, plus 5.6x end-to-end on GDP-Val with no measured degradation. For long agentic eval sweeps, that reframes inference speed as an experiment-iteration variable, not a UX nicety.
Each link below shares sources, entities, or timing with this story.
OpenAI launched the GPT-5.6 family on July 14: Sol (flagship), Terra (cost-optimized), and Luna (fast tier), live across ChatGPT, Codex, and the API the same day after a US-government-requested delay for security review. The numbers are loud. Sol scored 53.6 on Agents' Last Ex...
OpenAI shipped GPT-5.5 on April 23, six weeks after 5.4. The capability jump is real: 82.7% on Terminal-Bench 2.0 vs Claude Opus 4.7's 69.4%. The Pro tier nearly doubles Opus 4.7 on FrontierMath Tier 4 at 39.6% vs 22.9%. It uses 40% fewer tokens on Codex tasks while matching 5...
OpenAI released GPT-5.4 in Standard, Thinking, and Pro variants. Headline capabilities: native computer-use (75.0% on OSWorld-Verified, surpassing human 72.4%), 1M token context, and first-ever "compaction" support for longer agent trajectories. The Tool Search API is the buil...
Four frontier models. Five sealed engineering problems. The result everybody will quote is that Claude Fable 5 won. The result that should actually change how you work is buried three-quarters down the page. JuliaHub published an evaluation on July 30 running four frontier mod...
Alongside the July 29 launch of ChatGPT for Academic Researchers, free GPT-5.6 Sol Pro for 10,000 researchers this summer scaling to 100,000 through 2027 backed by over $250 million, OpenAI disclosed efficiency work on the harness underlying Codex and ChatGPT Work: 54% fewer o...
The July 10 refresh has Sol/Terra/Luna entering at 64.6/63.4/62.7% while Claude Fable 5 holds 80.3%, a +11.1 jump over Opus 4.8. Pro uses actively-maintained repos with no public ground-truth leakage, so its gap from the near-saturated Verified benchmark is the more honest sig...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.