Fetching from the wire…
Models2026-08-27 · source-backed
Published August 25, it's IBM's first family of dense decoder-only reasoning models, with the 30B flagship claiming state-of-the-art resolve rates on SWE-Bench Pro and Terminal-Bench. The 8B and 30B went through an agentic training curriculum on real sandboxes for software engineering, terminal operations and web research (Hugging Face). The r/LocalLLaMA thread reached 376 upvotes with top comments conceding Granite trails on benchmarks but valuing the permissive license. IBM also released a 470M Apache-2.0 ASR model claiming 12,600 RTFx on one H200.
Each link below shares sources, entities, or timing with this story.
For about a year, "run your agent locally" meant accepting a model that couldn't reliably call a tool twice in a row. That excuse is gone. Meta Superintelligence Labs published Muse Glimmer today: a 29.6B dense causal transformer, 52 layers, 6,656 hidden dim, with a ~1.8B ViT-...
DeepSeek released V4 Preview on April 24 with two open-weight variants: V4-Pro (1.6T total parameters, 49B activated via MoE) and V4-Flash (284B parameters, 13B activated). Both support 1M-token context windows. Both are Apache 2.0 licensed. Both are live right now on Hugging...
The walkthrough covers cluster provisioning, NVFP4 quantization and an OpenAI-compatible endpoint with reasoning support for the 95B-active MoE. Independent benchmark compilations put Qwen3.8-Max at 86.6 on Terminal-Bench 2.1 between GPT-5.6 Sol at 88.8 and Claude Opus 5 at 84...
granite-speech-5.0-470m-turboctc uses 16 conformer blocks trained with CTC on a 16,384 BPE head, temporal subsampling by 8, 128-frame block attention and self-conditioned CTC from the middle layer. IBM reports above 12,600 RTFx on a single H200, roughly 3.5 hours of audio per...
OpenAI shipped GPT-5.5 on April 23, six weeks after 5.4. The capability jump is real: 82.7% on Terminal-Bench 2.0 vs Claude Opus 4.7's 69.4%. The Pro tier nearly doubles Opus 4.7 on FrontierMath Tier 4 at 39.6% vs 22.9%. It uses 40% fewer tokens on Codex tasks while matching 5...
Three moves, two days, no coordination between them. August 10–11: GitHub shipped Ollama as a BYOK provider inside Copilot for JetBrains (GitHub Changelog). Unsloth released Unsloth Desktop with a command literally named unsloth start claude, which points Claude Code and Codex...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.