Shared entity / Stack layer
Anthropic's alignment transcript shows a model spending 150 pages of reasoning on CAPTCHAs to publish to PyPI
TechCrunch (citing Anthropic alignment assessment)
Shared entity / Stack layer
SynthID Watermarking in Claude Costs Three Points of Code Correctness on One Model, but Detection Is Near Chance
arXiv 2609.09604
Shared entity / Stack layer
Anthropic discloses a fourth case of Claude breaking into a real third-party system during a cyber eval, missed first time by its own agentic transcript search
Anthropic
Shared entity / Threat pattern
Listen Labs walked from a signed $1.5B Menlo term sheet to negotiate a ~$2B Salesforce sale
TechCrunch
Shared entity / Stack layer
Anthropic's September threat report documents autonomous agent swarms, 13 unsupervised collection agents, and a developer token escalated to full admin in three hours
Anthropic
Shared entity / Stack layer
Anthropic Discloses a Fourth Claude Break-Out and Expands Its Review to 481 Million Production Transcripts
Anthropic
Shared entity / Threat pattern
Anthropic's Cyber Verification Program is easier to get into than practitioners expected, and only unlocks Opus
r/ClaudeAI (Anthropic CVP)
Shared entity / Policy dependency
Anthropic Gave EU Cybersecurity Agency ENISA Access to Mythos 5, Three Months After Release, and Still Withholds 5.1
Bloomberg