Fetching from the wire…
Public story · 2026-08-14 · high
Gemini cut prices, Writer cut agent spend, DeepSeek gave away its orchestration layer, and a new coding agent went ad-subsidized and free.
Why now: All five moves landed within the same 48-hour window ending August 14, per VentureBeat, Google's blog, DeepSeek, and Product Hunt.
Google cut Gemini 3.7 Flash prices 50% on August 13, kicking off a rush to compress AI agent costs, per Google's blog.
Four more companies followed within the next day, each attacking a different layer of the stack: the model, the harness, the runtime, and distribution. For builders running agents in production, the bill they've tracked as one line item now splits across five prices that are all moving at once.
Writer said the same day that its Palmyra X6 model cuts agent costs 52%, not through a cheaper model but through harness engineering, per VentureBeat.
DeepSeek open-sourced its Harness under MIT on August 13, per DeepSeek, freeing the orchestration layer that sits between model and agent.
Hoplite, a YC S26 startup, launched the same window to run 'software factories' for teams, per its Product Hunt listing. The pitch: heavy token spend without building the infrastructure yourself.
Freebuff shipped an ad-subsidized coding agent priced at $0 on August 14, undercutting Cursor's $20 and $200 tiers, per its Product Hunt listing.
Each link below shares sources, entities, or timing with this story.
Gemini 3.7 Flash launched at a 50% introductory cut. Writer claimed 52% agent cost reduction via harness engineering. DeepSeek open-sourced its Harness under MIT so the orchestration layer is free. Hoplite (YC S26) launched to run "software factories" for teams that want to to...
Meituan MIT-licensed LongCat-2.0, a 1.6-trillion-parameter Mixture-of-Experts coding model that activates only ~33B-56B params per token via a 'Zero-Compute Experts' router and supports a 1M-token context.
Customers pay Salesforce per API call for CRM access, then sign a second contract with Anthropic for inference.
Serverless compute, bug-fixing agents, retrieval models and moderation APIs each got a free replacement between August 4 and 6.
The 284B-parameter model claims a six-fold agent capability jump over the prior Flash tier while undercutting proprietary pricing by roughly 60 percent.
Only GitHub's move carries a date, June 2026, while the other three don't say when they switched.
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.