Fetching from the wire…
Top 5 · 2026-03-05 · source-backed
Junyang Lin, Alibaba's lead Qwen researcher and one of their youngest P10 employees, resigned along with key team members Binyuan Hui (Coder series), Bowen Yu (Instruct/post-training), and Kaixin Li. The departures came weeks after Qwen 3.5's release — models practitioners call "the most capable agentic coding model at that size." Root cause: tension between the research team and Alibaba's product division pushing DAU metrics for the Qwen App over research priorities. This threatens one of the strongest open-weight model families at a critical moment. Action: If you depend on Qwen models, continue using them but diversify your model strategy. Watch for where the departing researchers land. Simon Willison | HN 734pts
Each link below shares sources, entities, or timing with this story.
Alibaba's model was reported best overall on Artificial Analysis' Agentic Index, drawing 540 points on HN. Readers watching the page saw Qwen at 55.4 vs Opus Max at 55.3, then on reload Opus Max at 59.2 vs Qwen at 58.4. George from Artificial Analysis replied in-thread that th...
Junyang Lin (tech lead who built Qwen from lab project to 600M+ downloads) and Yu Bowen (post-training head) resigned one day after Qwen 3.5 launched. Huibin (Qwen Code lead) had already left for Meta in January. The catalyst: Alibaba dismantled Lin's vertically-integrated R&D...
Alibaba released Qwen3.6-27B on April 22. Dense architecture. Open weights. 77.2% on SWE-bench Verified, within 3.7 points of Claude Opus 4.6. On SkillsBench, it scores 48.2% versus its own 397B MoE predecessor's 30.0%. That's a 77% improvement with 14.8x fewer parameters. Let...
Somebody diffed the configs. Zero architectural changes. Same 64 layers, same 5,120 hidden dimension, same hybrid Gated DeltaNet → FFN / Gated Attention → FFN block structure as Qwen3.6-27B. The r/LocalLLaMA post showing this hit 945 upvotes and 157 comments, and Hugging Face...
The open-weight race just changed constraint. Moonshot AI suspended all new consumer subscriptions on July 20, roughly 48 hours after Kimi K3 launched, because request volume pushed its compute cluster to capacity. Remaining GPUs are reserved for existing paid subscribers. Tec...
25,000 fake accounts. 28.8 million Claude conversations. Six weeks. And the thing they were harvesting wasn't trivia, it was software engineering and agentic reasoning. In a June 24 letter to US senators and the White House, Anthropic alleged that operators tied to Alibaba's Q...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.