Fetching from the wire…
Top 5 · 2026-05-19 · source-backed
Your AI coding budget just got a lot harder to predict.
A viral analysis on Hacker News (413 points, 396 comments) makes the case that every major AI lab has been running a loss-leader program, and the correction is starting. Two concrete dates matter. GitHub transitions all Copilot plans to token-based AI Credits on June 1: Pro ($10/mo) gets $15 in combined credits, Pro+ ($39/mo) gets $70, and a new Max tier ($100/mo) gets $200. Anthropic splits Claude subscriptions into usage pools effective June 15. Code completions and next-edit suggestions stay unlimited and free, but everything agentic now has a meter running.
The catalyzing number: Uber reportedly burned through its entire 2026 AI budget by April. Let that sink in. A company with massive engineering resources couldn't forecast agentic compute costs accurately enough to make their annual budget last four months.
This isn't surprising if you've been running agentic workloads. A chat completion is a few thousand tokens. An agent loop that reads files, writes code, runs tests, reads errors, and iterates can burn 100K+ tokens per task. I've watched single Claude Code sessions in my personal projects consume what used to be a day's worth of API credits in 20 minutes. Flat-rate pricing was never going to survive that math.
Windsurf is making the same move. They bumped Pro from $15 to $20/month, added a $200/month Max tier, and shifted Bugbot to usage-based billing effective June 8. The bundling of Devin Cloud softens the price increase, but the direction is unmistakable. Every tool is migrating from "all you can eat" to "pay for what you consume."
The builder move: audit your agentic workload costs this week, before the switchover dates. Measure actual token consumption per task type. Budget for 3-5x what chat-based usage costs. And look seriously at the custom model story below, because the cost pressure is exactly why platforms are training their own models.
Each link below shares sources, entities, or timing with this story.
43.3% on Frontier-Bench v0.1. Opus 4.8 scored 18.7%. That's not an incremental bump, that's the same benchmark with a different shape of answer. Anthropic released Claude Opus 5 on July 24 at $5/$25 per million input/output tokens, exactly half of Fable 5's $10/$50, while matc...
Here's the counterweight to the PMF story. Microsoft began revoking internal Claude Code licenses for most employees on May 14 with a June 30 deadline. The Experiences and Devices division, the team behind Windows, Microsoft 365, Outlook, Teams, and Surface, is moving engineer...
GitHub paused all new individual plan sign-ups and gutted the model lineup. The reason they gave is the quiet part said loud: agentic workflows consume far more resources than the plan structure can support. Starting April 20, new sign-ups for Copilot Pro ($10/mo), Pro+ ($39/m...
GitHub announced that every Copilot plan moves to usage-based billing via "AI Credits" on June 1, 2026. The flat-rate model is dead. Every plan gets a monthly credit allotment calculated by actual token consumption (input, output, cached) at listed API rates. Code completions...
Alibaba released Qwen3.6-27B on April 22. Dense architecture. Open weights. 77.2% on SWE-bench Verified, within 3.7 points of Claude Opus 4.6. On SkillsBench, it scores 48.2% versus its own 397B MoE predecessor's 30.0%. That's a 77% improvement with 14.8x fewer parameters. Let...
GitHub quietly updated its Copilot pricing multiplier table, and the numbers are jarring. Starting June 1, 2026, every Copilot interaction gets priced in "AI Credits" with per-model multipliers: Claude Opus at 27x the base rate, Claude Sonnet at 9x, and base completions at 1x....
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.