Fetching from the wire…
Models2026-06-10 · source-backed
Open-weight trackers put it at 77.2% on MMLU-Pro, surpassing the prior-gen 27B model. If it holds, it's another small open model matching a previous flagship and lowering the bar for on-prem deployment. Single-source benchmark, so verify the numbers before you build on the exact figure. But the direction is the same one Cohere just reinforced: smaller open weights keep eating last year's frontier.
Each link below shares sources, entities, or timing with this story.
Google is closing the current Imagen generation's API endpoints and pointing callers at the Gemini image model (LLM Stats). It consolidates image generation under a single Gemini surface, which means anyone still pinned to Imagen endpoints has a migration today. Confirmed thro...
The models are good. The license is the real story. Google released Gemma 4 on April 2 with four variants: E2B, E4B, 26B MoE, and 31B Dense. All built on the Gemini 3 architecture. The 31B Dense variant claimed #3 on Arena AI's text leaderboard, beating models 20x its size. Th...
North Small Translate is a mixture-of-experts model with 218B total and 25B active parameters, covering 50 languages. It scores 83.60 on WMT26, and an agentic multi-pass variant scores 84.36. DeepL NextGen scores 81.37 and Google Translate 68.20. On book-length translation in...
Google DeepMind released Gemma 4 on April 2 with four model sizes (E2B, E4B, 26B MoE, 31B Dense) under Apache 2.0. Multimodal (text, vision, audio). 256K context. Native thinking and tool-calling optimized for agentic workflows. Day-zero ecosystem support across vLLM, llama.cp...
Google's HF org lists diffusiongemma-26B-A4B-it (~4B active), an image-text-to-text Gemma member that's diffusion-style rather than purely autoregressive (Hugging Face). No detailed announcement yet, which is why I'm flagging it low. But a diffusion approach inside the Gemma o...
GLM-5.1 from Zhipu AI scored 58.4% on SWE-bench Pro. GPT-5.4 scored 57.7%. Claude Opus 4.6 scored 57.3%. That's the first time an open-weight model has ever topped a major coding benchmark against the best proprietary models. The specs matter. GLM-5.1 is a 754B-parameter mixtu...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.