DeskQwen 3.5 397B MoE LaunchWweb·high signalXBlueskyLinkedInCopy linkAlibaba launched Qwen 3.5: 397B total/17B active MoE, Apache 2.0, 1M context, 201 languages, SWE-bench 76.4↳ Follow the threadShared entity / Policy dependencyA Fine-Tuned 4B Qwen in 2.6 GB Beats GPT-5.6 on a Transit-Kiosk Agent Benchmark, and PEFT Gains Vanish by 27BarXiv 2609.10016Shared entity / Stack layerAlibaba Is Leading a $300M Round in Model-Testing Startup UniPat AI at a $2.5B ValuationTechmeme (via Bloomberg)Shared entity / Threat patternAnthropic ties about 200 million Claude exchanges to distillation by Alibaba, Moonshot and DeepSeek, and says Moonshot passed Claude answers off as KimiTechCrunch (corroborated by CNBC, SCMP, Anthropic threat report; via r/singularity)Stack layer / ContrastEdge0 runs a 35B MoE on Apple Silicon in 2.9 GB of active memory by streaming experts off SSDGitHubPolicy dependency / Stack layerCROSS-CATEGORY: Three Independent Agent-Action Gates Shipped in 48 Hours, All Judging the Command Against Stated IntentProduct Hunt, github.com/AGGIB/Stroq and rewarelabs.com (three independent sources; the 72% figure is Reware's own)Stack layer / ContrastCohere Released an Open-Weight 218B Translation Model That Beats DeepL and Google Translate on WMT26Hugging Face (Cohere Labs)Stack layer / Update threadAWS published a working recipe for self-hosting the 2.4-trillion-parameter Qwen3.8 on HyperPod with vLLM and NVFP4AWS Machine Learning BlogStack layer / Update threadTransformers 5.17.0 ships a 780B-parameter MoE with a 1M context and a per-channel-gated linear attention modelGitHub