Fetching from the wire…
Top 5 · 2026-05-02 · source-backed
69% of all input tokens in production LLM traces are system prompts. Let that sink in for a second.
Datadog's State of AI Engineering 2026 dropped yesterday, and it's the best empirical snapshot we have of how companies actually use LLMs. Not how they demo them. Not how they pitch them to investors. How they run them in prod, at scale, across thousands of deployments.
The numbers that matter: 5% of all LLM spans report errors, with 60% of those coming from rate limits. Framework adoption (LangChain, Pydantic AI, Vercel AI SDK) doubled from 9% to 18% of organizations year-over-year. Anthropic Claude grew 23 percentage points of provider share while OpenAI maintains 63%. And 69% of companies now use 3+ models in production.
But that 69% system prompt stat is the one I keep coming back to. If you're optimizing for cost and latency, system prompts are where the money is burning. Two-thirds of your input tokens aren't user queries or retrieved context. They're the instructions you wrote once and send every single call. This means prompt compression, caching strategies, and system prompt architecture aren't premature optimization. They're the first thing you should look at.
The framework adoption doubling is interesting because it tells us the "just use the raw API" phase is ending for most teams. When you're running 3+ models with error handling, fallbacks, and observability, you need an abstraction layer. The question is which one wins. LangChain's been losing mindshare to lighter alternatives, but Datadog's data shows it's still growing in absolute terms.
The Anthropic market share jump (23pp) is the competitive signal here. OpenAI still dominates at 63%, but that dominance is eroding fast. A year ago it was closer to 80%. Claude is eating into that gap primarily through coding and agent use cases, which is exactly what you'd expect given Claude Code's adoption curve.
What to do about it: Audit your system prompts today. If you're sending 2,000+ tokens of instructions on every call, look at Anthropic's prompt caching (which gives you 90% off cached input tokens) or restructure your prompts to minimize repetition. The Datadog data says this is where most production spend actually lives.
Each link below shares sources, entities, or timing with this story.
1. Flip your multi-model pipeline to review-then-generate. Instead of using a reasoning model to plan before code generation, let the specialist generate freely and use reasoning tokens for review. Paper shows 90.2% pass@1 vs 87.2% for the planning pattern. Source 2. Audit you...
Five months ago, Anthropic was running at $9B annualized. Today it's $30B. CNBC named them #1 on the 2026 Disruptor 50, above OpenAI for the first time. The numbers from Daniela Amodei's interview are hard to process. $1B run rate in December 2024. $9B end of 2025. $14B Februa...
The person who coined "vibe coding" and co-founded OpenAI just chose Anthropic. TechCrunch confirmed on May 19 that Andrej Karpathy has joined Anthropic's pre-training team, where he'll start a new group using Claude to accelerate pre-training research under team lead Nick Jos...
The IDE market is fragmenting, and this week drew the sharpest lines yet. Cursor 3 launched as a rebuilt agent-orchestration platform in Rust and TypeScript, replacing the VS Code fork with an Agents Window for dispatching and monitoring multiple AI coding agents. Anysphere hi...
Twelve months ago, OpenAI led Anthropic by 41 points in enterprise adoption. Today that gap is 8. Enterprise Technology Research's survey of roughly 500 respondents shows OpenAI dropping from 62% adoption (September 2025) to 56% (March 2026) while Anthropic surged from 21% to...
The RSI debate has been vibes and timelines for two years. This week a frontier lab published an actual measurement from inside its own walls. The Anthropic Institute reported an 8x increase in lines of code merged into its codebase in 2026 versus the 2021–2024 baseline. The t...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.