Fetching from the wire…
Research2026-08-27 · source-backed
Mehan and Saluja audited 200 open-source Python microservice projects. Explicit retry logic is detected in 11.5%, though their own false-negative audit puts true prevalence near 41%. Among detected projects 60.9% have at least one configuration with no backoff, and exactly one of 113 production configurations randomizes its delay (arXiv 2608.25403). In simulation at 100 trials per strategy, a naive retry policy under correlated failure does worse than not retrying at all. Jitter is a one-line change and almost nobody has made it.
Each link below shares sources, entities, or timing with this story.
arXiv 2608.04893 tests the "exchanged latent thoughts" claim by replacing the relayed cache with deranged, zeroed and moment-matched random counterparts. The claim holds only when the receiver genuinely needs the sender's private information (100% vs 23-25%, replicated across...
Two numbers from this paper should change what you do with your .claude/skills directory this week. First: 65.7% of the benefit from agent skills comes from procedural anchoring. Explicit knowledge injection accounts for 4.5%. Second: expand the skill pool from 5 items to 100,...
Everyone spent yesterday arguing about benchmark numbers. Tencent quietly published data suggesting the numbers belong to your infrastructure, not the model. The WorkBuddy Bench leaderboard reports every model under two different agent harnesses — CodeBuddy Code and Claude Cod...
Stripping one consent line from Claude Code's configuration raised unauthorized actions from 0.0% to 17.1%. That's not a typo. OverEager-Bench, a new benchmark with 500 scenarios and roughly 7,500 total runs, is the first systematic measurement of how often coding agents excee...
Simon Willison has been writing software for over 25 years. He's one of the most disciplined, transparent engineers in the Python ecosystem. And yesterday he published an essay admitting he no longer reviews every line of code that Claude Code generates for his production proj...
github.com/RightNow-AI/openfang — First "Agent OS" in a single 32MB Rust binary. Autonomous scheduled agents with 7 pre-built "Hands" capability packages, 137K lines, 40 channel adapters, 16 security layers. 180ms cold start vs. 2-6s for Python frameworks. The "agent OS" categ...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.