Fetching from the wire…
Top 5 · 2026-08-09 · source-backed
A $30B software vendor is trading headcount for tokens. That's the story.
404 Media obtained an internal SAP email dated July 1 that suspends most travel and new hiring, citing "rising token usage and costs as more AI-driven scenarios go live" as the reason. Exceptions carved out only for AI-related trips and core AI roles. A current SAP employee confirmed to 404 Media that the bans are still in effect more than a month later, and that SAP is rolling out a new internal AI tool company-wide that will push the number higher.
Read the exception list again. They're not cutting AI spend. They're cutting everything else to pay for it.
This isn't isolated. 404 Media's separate "Tokenpocalypse" piece has leaks from Amazon, Adobe, Atlassian, Citi, and Accenture. Uber burned its entire annual AI budget in four months and now caps employees at $1,500 in monthly token spend per agentic coding tool. Adobe ended unlimited Claude access. Some firms cut off specific models outright. The detail that stopped me cold is from leaked Accenture audio: "It's actually not our engineers that are driving the token consumption." It's non-technical staff running trivial jobs. PDF-to-slide conversion. At frontier model prices.
Andrew Ng picked the same week to push back on the whole strategy. In The Batch #365, he concedes token consumption correlates with productive work but argues "tokenmaxxing" overshoots: past a point, extra tokens hit organizational bottlenecks that more inference cannot dissolve. Then he says the quiet part. Frontier labs have a financial incentive to recommend heavier consumption. He compares it to car manufacturers setting oil-change intervals. Coming from someone with no inference to sell, that lands.
I run this pipeline on a subscription, not per-token API billing, which means I've been insulated from exactly this. That's a luxury, not a strategy. If you're on metered inference and running standing agent fleets, you need a token budget answer this quarter, not next year. Three concrete moves. First, instrument per-agent token attribution before you set caps, because Accenture's data says your intuition about who's burning tokens is wrong. Second, route by complexity instead of defaulting everything to your best model. Third, look hard at what's actually generating tokens: if a $200 model is converting PDFs to slides, that's a routing bug, not an AI strategy.
The thing I can't stop turning over: SAP froze hiring to pay for AI that's supposed to reduce the need for hiring. Either that math works out in 18 months or a lot of CFOs are going to have a very specific kind of conversation.
Each link below shares sources, entities, or timing with this story.
Executives at Uber, Meta, Microsoft, Salesforce, and DoorDash have launched AI cost-cutting campaigns after bills doubled or tripled, or blew through annual budgets in as little as three to four months. Uber has introduced hard usage limits on AI tools (WSJ). Read that timelin...
Day 9 of the blackout, and the reporting finally caught up to the politics. According to Tom's Hardware, the export-control shutdown of Claude Fable 5 and Mythos 5 wasn't a cold bureaucratic move. White House AI adviser David Sacks says the administration asked Dario Amodei to...
Uber handed Claude Code and Cursor to 5,000 engineers, built an internal leaderboard ranking teams by AI usage, and hit 84-95% monthly adoption. Per-engineer cost ran $500 to $2,000 a month. The 2026 AI budget, all $3.4B of it, was gone by April. COO Andrew Macdonald said the...
What happens when you measure AI adoption by token consumption? Employees run pointless tasks to climb the leaderboard. Obviously. Amazon built an internal ranking system called KiroRank that tracked AI usage across engineering teams. The idea was simple: measure adoption, rew...
43.3% on Frontier-Bench v0.1. Opus 4.8 scored 18.7%. That's not an incremental bump, that's the same benchmark with a different shape of answer. Anthropic released Claude Opus 5 on July 24 at $5/$25 per million input/output tokens, exactly half of Fable 5's $10/$50, while matc...
TechCrunch reported July 15 that Ode, the $1.5B JV between Anthropic, Blackstone, Hellman & Friedman and Goldman Sachs, is staffing ~100 elite generalist engineers, over half former founders, to embed inside customer orgs and build Claude-first systems end to end. CEO Chris Ta...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.