Fetching from the wire…
Markets2026-08-21 · source-backed
Poolside told investors the license is non-exclusive, covering the system used to train its Laguna open model, plus job offers to 109 employees and a separate $1B investment at a $12B pre-money valuation. The letter insists this is neither acquisition nor acquihire, and the three cofounders stay. r/LocalLLaMA It follows the structure Nvidia used with Groq ($20B) and Enfabrica ($900M), and it makes the largest GPU vendor a direct competitor to its own model-building customers. Reported by Newcomer and The Information; I'd treat the exact terms as investor-letter framing until confirmed.
Each link below shares sources, entities, or timing with this story.
PR #19378 landed in llama.cpp this week, and I think most people are underselling what it means. Backend-agnostic tensor parallelism via --split-mode tensor makes multi-GPU inference work across AMD, Intel, and Apple Silicon. Not just CUDA. Everything. For context: llama.cpp h...
3B active parameters, beats Qwen3.5-35B-A3B on AIME 2025 (92.4 vs 91.9), LiveCodeBench v6 (87.2 vs 74.6), and surpasses the larger Nemotron-3-Super-120B. Available on Ollama and HuggingFace under open license. Source
Alibaba published a fine-grained MoE with 2.4T total / 95B active, 512 experts, and a 92-layer hybrid full/linear attention backbone. vLLM shipped day-0 support verified on NVIDIA and AMD with ready 4-bit checkpoints (NVFP4 at 1.32 TiB for an 8xB300 node, MXFP4 at 1.45 TiB for...
Poolside's model (118B total / 8B active, 1M context, 70.2% on Terminal-Bench 2.1, NVFP4 weights fitting a single DGX Spark) is getting its first sustained hands-on scrutiny and the verdict splits in a specific way. One user reported it solved a data-restructuring problem that...
Announced July 27 with Microsoft, IBM, Red Hat, Palantir, CrowdStrike, Cloudflare, Databricks, Hugging Face, LangChain, Nous Research, Reflection AI, Thinking Machines Lab, SpaceXAI and the Linux Foundation. Huang's framing is pointed: during the Hugging Face incident "closed...
NVIDIA's freshly released 4B entrant failed all custom agentic benchmarks where Qwen 3.5 4B Q8 passed every one. First head-to-head from GTC model releases. 142 upvotes on r/LocalLLaMA. Source
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.