Fetching from the wire…
Public story · 2026-08-16 · high
The 27.78B-parameter model scores 61.7% on SWE-Bench Pro and runs on a single GPU, per AI Release Tracker.
Why now: Alibaba released Qwen3.8-27B on August 14, covered in the August 16 briefing alongside Qwen3.8-Max's launch 11 days earlier.
Alibaba's Tongyi Lab released Qwen3.8-27B on August 14, a 27.78B-parameter dense model scoring 61.7% on SWE-Bench Pro, per AI Release Tracker.
Apache 2.0 licensing lets any team fine-tune and redistribute the model without negotiating access, a bigger deal than the score for teams building on it.
LiveCodeBench hits 90.3%, GPQA Diamond hits 89.2%, and Terminal-Bench 2.1 climbs to 73.0 from 63.4 in the prior Qwen3.6-27B release. DeepSWE 1.1 jumps from 13.3 to 42.2, the sharpest gain in the set.
The model takes text, image, and video input. Its native 262,144-token context stretches to 1M through YaRN.
It arrived 11 days after Qwen3.8-Max, Alibaba's other Qwen3.8 release. AI Release Tracker doesn't say whether Max carries the same Apache 2.0 terms.
Qwen3.8-27B, not Qwen3.8-Max, is the model most teams will actually fine-tune, because it's the one that fits on a single GPU. Watch fine-tune and download counts over the next few weeks. If Qwen3.8-27B outpaces Max there, deployability beat benchmark size again.
Each link below shares sources, entities, or timing with this story.
Moonshot AI dropped Kimi K2.6 today and the numbers are hard to ignore. One trillion parameters total, 32 billion active per token across 384 experts, 256K context window, and native multimodal input. It scores 58.6 on SWE-Bench Pro versus GPT-5.4's 57.7 and Claude Opus 4.6's...
For about a year, "run your agent locally" meant accepting a model that couldn't reliably call a tool twice in a row. That excuse is gone. Meta Superintelligence Labs published Muse Glimmer today: a 29.6B dense causal transformer, 52 layers, 6,656 hidden dim, with a ~1.8B ViT-...
Alibaba announced it August 3: sparse MoE with ~95B active per token, 1M context, 128k max output, $2/M input and $6/M output with $0.25/M cached. 67.4 on Terminal-Bench 2.1 (up from 61.0 for 3.7 Max), #4 on Frontend Code Arena at 1,668 Elo, #2 on Vals Index among open-weight...
Somebody diffed the configs. Zero architectural changes. Same 64 layers, same 5,120 hidden dimension, same hybrid Gated DeltaNet → FFN / Gated Attention → FFN block structure as Qwen3.6-27B. The r/LocalLLaMA post showing this hit 945 upvotes and 157 comments, and Hugging Face...
Moonshot AI dropped Kimi K2.7-Code on Hugging Face on June 12. The specs are loud: 1T-parameter MoE with 32B active across 384 experts, a 256K context window, Modified MIT license, tuned for long-horizon agentic software engineering (MarkTechPost). Moonshot reports +21.8% on K...
Alibaba released Qwen3.6-27B on April 22. Dense architecture. Open weights. 77.2% on SWE-bench Verified, within 3.7 points of Claude Opus 4.6. On SkillsBench, it scores 48.2% versus its own 397B MoE predecessor's 30.0%. That's a 77% improvement with 14.8x fewer parameters. Let...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.