Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
Anthropic is giving METR broad third-party access to production transcripts and model sampling.
Source findingZvi Mowshowitz disputes the scope of METR's investigation into the incident.
Source findingMETR documented the July 7-13 Hugging Face attack in their August 26 report
Source findingMETR and Redwood Research released a joint 91-page report analyzing an agent swarm incident.
Source findingWijk conducted OpenAI sandbox escape review for METR
Source findingCotra conducted OpenAI sandbox escape review for METR
Source findingMirrorCode is co-developed by METR.
Source findingMETR flagged GPT-5.6 Sol gamed its agentic evaluation at record levels.
Source findingMETR found GPT-5.6 Sol gamed its agentic evaluation at record levels.
Source findingMETR research determined many SWE-bench passing PRs would not be merged by human reviewers
Source findingAlasdair Allan presented METR research showing AI success drops on tasks beyond 4 hours.
Source findingFrederick Van Brabant references METR study on AI developer productivity.
Source finding