Fetching from the wire…
Public story · 2026-08-15 · high
It holds 262K tokens of context and reportedly beats Alibaba's larger Qwen3.7-Plus on coding and office work.
Why now: Alibaba posted the weights on August 14, 2026, and The Decoder's coverage landed in the news cycle the next day.
Alibaba's Qwen team released the weights for Qwen3.8-27B on Hugging Face and ModelScope on August 14, 2026, per The Decoder.
The Apache 2.0 license lets any company deploy Qwen3.8-27B commercially, no fee, no legal review. That's a change from Alibaba's recent Qwen releases, which shipped closed, putting a 27.78B multimodal model within reach of teams that can't negotiate enterprise terms.
Qwen3.8-27B is a dense multimodal model with 27.78 billion parameters, small enough to run on a single high-end local GPU. It handles text, image, and video, with native context at 262,144 tokens and room to extend toward 1 million. The architecture pairs Gated DeltaNet with Gated Attention.
Per The Decoder, Qwen3.8-27B reportedly beats Qwen3.7-Plus, a much larger Alibaba model, on coding and office tasks. The source doesn't say by how much or which benchmarks it used, so treat this as the release's own claim, not an independent test.
This release lands amid a run of Chinese frontier-model launches, with related coverage counting four separate releases in under a month.
Watch whether Alibaba keeps this license on its next model. Qwen3.7-Plus shipped closed, so this open release can be the exception rather than a new default.
Each link below shares sources, entities, or timing with this story.
Somebody diffed the configs. Zero architectural changes. Same 64 layers, same 5,120 hidden dimension, same hybrid Gated DeltaNet → FFN / Gated Attention → FFN block structure as Qwen3.6-27B. The r/LocalLLaMA post showing this hit 945 upvotes and 157 comments, and Hugging Face...
Pricing runs on four model tiers instead of per seat, and the platform slots into DingTalk's 20 million business accounts.
Any operator clearing 20 million dollars in revenue over any rolling 12 months must sign a separate deal with Moonshot.
Launched August 8 as the new Quality Mode at grok.com/imagine and in the Grok mobile apps, pitching precision editing, crisp text rendering, and improved factuality, with API access promised but not shipped (The Decoder). On the August 7 Arena leaderboards the faster "low" var...
The number that reframes everything isn't ten. It's two thousand. OpenAI published "Ten advances in mathematics and theoretical computer science" on August 1, claiming an internal version of Astra produced new results on ten problems that had seen no progress on the main resul...
The system card reports browser-agent injection falling from 31.5% to 3.70% on the model alone, then to 0% with Auto Mode enabled, where one layer scans incoming data for hidden instructions and a second blocks dangerous actions before execution. Gray Swan's independent genera...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.