Stack layer / Contrast
Controlled agentic CAD comparison: six unattended runs, 16 failures, 9 of which the tool never reported
ModelRift
Policy dependency / Stack layer
Claude Code's agent view adds a peek panel so you can answer a blocked session without leaving the list
Claude Code Docs
Policy dependency / Stack layer
A replay of 68,266 real Claude Code requests says plain LRU beats the clever KV-cache policies
GitHub
Stack layer / Threat pattern
Claude Code 2.1.269 Ships a Plugin Eval Runner and a Knob to Raise the Workflow Tool's Concurrent Agent Cap to 256
Anthropic (claude-code CHANGELOG)
Stack layer / Contrast
Claude Code ships `claude plugin eval`: every case runs with and without your plugin, and the delta is the only score that proves it did anything
Claude Code docs
Policy dependency / Stack layer
SGLang Hit With Unauthenticated Pickle RCE via /update_weights_from_tensor, the Fourth Critical Inference-Stack CVE in Four Weeks
CERT Coordination Center
Stack layer
AgentCore Evaluations ships 16 evaluators for the failure mode infrastructure monitoring cannot see
AWS Machine Learning Blog
Contrast / Follow-up thread
r/MachineLearning asks whether ML publishing is past the point of no return at ~447 cs.LG papers in a day
r/MachineLearning (arXiv cs.LG listing checked directly)