NewsMicrosoft GRP-Obliteration Breaks Safety Across 15 LLMsMicrosoft Security Blog·high signalXBlueskyLinkedInCopy linkSingle prompt jailbreak unaligns 15 models across 6 families. GPT-OSS-20B attack success from 13% to 93%. Also breaks diffusion models.SourceSource pageMicrosoft Security Blog↳ Follow the threadPolicy dependency / Stack layer'Do this as quickly as possible' repeatedly got a Claude session flagged by a corporate security directorr/ClaudeAIPolicy dependency / Stack layerPython's Import Statement Is an Execution Boundary: 90% of Initialization-Activated Advisory Vulnerabilities Are High or CriticalarXiv 2609.14791Policy dependency / Threat patternA Fine-Tuned RoBERTa-Large Permission Gate Matches Claude Haiku 4.5 at Deciding What an Agent May ToucharXiv 2609.15422Stack layer / ContrastEmergence World ran 10 agents per world for 16 days and found no frontier model contained an injected attack — one acted on poisoned memory 46 hours laterarXivStack layer / ContrastOPEN-1B Claims Bitwise-Reproducible Training Across Heterogeneous Commodity HardwarearXiv 2609.17380Stack layer / ContrastByteShape's Qwen3.8-27B quants hit 99.63% of BF16 at 3.84 bpw, and argue KL divergence is the wrong quant metricByteShape (via r/LocalLLaMA, 126 upvotes)Stack layer / ContrastTypeSafe AI ships Jev, a model that returns typed probabilistic values instead of text, at $0.042 per million input tokens and free outputTypeSafe AIStack layer / Threat patternMemRiskBench Scores Long-Horizon Agent Memory Risks Deterministically, With No LLM Judge on the Pass/Fail PatharXiv 2609.14976