SourcesMoltbook Safety Paper: Alignment Vanishes in Self-Evolving AI SocietiesHuggingFace Daily Papers·medium signalXBlueskyLinkedInCopy link165 HF upvotes. Safety alignment degrades when agents interact in social networks. Safe individual agents produce unsafe collective behavior.SourceSource pageHuggingFace Daily Papers↳ Follow the threadStack layer / Contrastbartowski measured which tensors actually break under quantization, and Q3_K_M got 15% smaller for itHugging Face / bartowski (via r/LocalLLaMA)Policy dependency / Stack layerAltman tells OpenAI staff the company is open to slowing frontier development and wants other labs to matchReuters / Bloomberg (via r/singularity)Stack layer / Threat patternA Malicious Super-App Can Silently Own Every Mini-App Inside It, and Russia's MAX Demonstrates the Full SetarXiv 2609.11814Stack layer / ContrastEvoSafeHarness searches policies and code together to build a per-model safety harness, cutting attack success from 45.6% to 10.0%arXiv (2609.05903)Stack layer / Threat patternClaude Code 2.1.269 Ships a Plugin Eval Runner and a Knob to Raise the Workflow Tool's Concurrent Agent Cap to 256Anthropic (claude-code CHANGELOG)Stack layer / ContrastPoland's Blik Ran Its First Agentic Payment, Where a User Pre-Authorized an Agent to Buy an Item When It Came Back in StockFinextraPolicy dependency / Stack layerA replay of 68,266 real Claude Code requests says plain LRU beats the clever KV-cache policiesGitHubStack layer / Follow-up threadBenchmark Radar Ships a Daily-Updated Catalog of 1,283 AI Benchmarks With 12,916 Numeric Score ObservationsarXiv 2609.11115