Fetching from the wire…
Tools2026-08-14 · source-backed
Willison's August 13 release adds Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite plus gemini-embedding-2 and -001, rebuilding on LLM 0.32's structured message and streaming APIs so reasoning, tool calls and results emit as typed stream events while preserving Gemini thought signatures. Google Search, URL context and code execution now run through -T GoogleSearch, -T URLContext and -T CodeExecution, combinable with local function tools in one call for Gemini 3 models. The 35 removed model IDs are a quiet signal about how much churn Google's catalog is generating.
Each link below shares sources, entities, or timing with this story.
Gemini 3.6 Flash launched July 21 at $1.50/1M in, $7.50/1M out, claiming 17% fewer output tokens than 3.5 Flash, DeepSWE code precision up from 37% to 49%, OSWorld-Verified computer use at 83% (from 78.4%), and a knowledge cutoff finally moved from January 2025 to March 2026....
Announced July 30, Gemini 3.1 Flash-Lite and 3.5 Flash join Cohere and Meta options, with Oracle explicitly framing model selection as per-scenario price-performance. The incumbent ERP vendor is conceding the model layer entirely and defending the data and workflow layer. That...
Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber, the last purpose-built to find and patch vulnerabilities and pitched as a cheap alternative to large security-specialized models like Mythos (DeepMind). Splitting a cheap tier into a security SKU is new packa...
Hardcoded model IDs now return API failures, and teams moving to 2.5-flash face a second mandatory cutover by October 16 (Google). Three major versions in under a year. The lesson isn't which model to pick, it's to stop hardcoding model IDs entirely. Abstract the selection or...
Google's Gemini 3.2 Flash appeared in the Gemini iOS app and AI Studio before any official announcement. It showed up on LM Arena benchmarks. And the numbers are real: 92% of GPT-5.5's coding and reasoning performance with sub-200ms latency at roughly 1/15th the cost. Source:...
Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image) generates images in as little as four seconds and is live in AI Studio, the Gemini API, AI Mode, and the Gemini app. Gemini Omni Flash entered public preview for video generation and conversational editing at $0.10 per second, m...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.