Fetching from the wire…
Research2026-08-09 · source-backed
arXiv 2608.01326 models compaction as two games: a Context Selection Game (retain a subset) and a Context Generation Game (summarize into a bounded message). It proves the generation game is equivalent to one-way communication complexity, so the minimum compaction budget for answering a query set within a target error is exactly that complexity. Then it proves a strict separation: generation-based compaction outperforms selection-based for certain query classes. Summarize-then-drop beats retrieve-and-keep on a provable set of workloads. They demonstrate by measuring Anthropic's own compaction endpoint on set-membership queries, which gives you a way to score a deployed algorithm against the optimum instead of eyeballing whether it "feels lossy."
Each link below shares sources, entities, or timing with this story.
Context engineering just got its cookbook. Anthropic's Claude Cookbook published a complete guide to three API-level primitives that, used together, reduced a research agent's peak context from 335K tokens to a sustained 50-80K range. Server-side compaction. Tool-result cleari...
1. Flip your multi-model pipeline to review-then-generate. Instead of using a reasoning model to plan before code generation, let the specialist generate freely and use reasoning tokens for review. Paper shows 90.2% pass@1 vs 87.2% for the planning pattern. Source 2. Audit you...
Stanford's Denisov-Blanch group built a maturity model for AI adoption scored entirely from artifacts already in version control, applied it to 441 repositories, and found something I've been assuming without evidence. RAMP is a four-level model derived only from committed AI...
For two years the technique was accumulation. Longer system prompts, longer CLAUDE.md, more numbered do/don't lists, more "always verify your work" imperatives. Anthropic's context-engineering guidance for Claude 5 models inverts it, with an 80% deletion figure attached. The s...
The Anthropic saga escalated from policy dispute to existential test this week. Amodei told a Morgan Stanley conference Anthropic has "no choice" but to challenge the Pentagon's supply chain risk designation in court — the first time a US tech company has ever received this la...
Anthropic ran a de novo binder campaign where Claude researched each target's biology, picked docking sites, installed open-source tools from their public repos itself, and composed 24 workflows with no human making a design decision. Of 1,320 designs synthesized and measured...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.