Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
Tool Search Tool improves Opus 4.5 tool-selection accuracy from 79.5% to 88.1%
Source findingOpus 4.5 leads SWE-bench Feb 2026 at 80.9% accuracy.
Source findingOpus 4.5 is built by Anthropic.
Source findingTool Search Tool improved selection accuracy from 79.5% to 88.1% on Opus 4.5
Source findingOpus 4.7 is a clear intelligence upgrade over Opus 4.5.
Source findingGemini 3.6 Flash competes with Anthropic's Opus 4.5 in the frontier model market.
Source findingClaude Sonnet 4.6 achieves Opus 4.5-comparable performance at Sonnet pricing
Source findingSemiAnalysis tracked benchmark performance gaps between Anthropic and OpenAI models across eras.
Source findingOpus 4.5 agents systematically out-negotiated Haiku 4.5 in the Project Deal experiment.
Source findingTool Search Tool improves Opus 4.5 tool-selection accuracy from 79.5% to 88.1%
Source findingTool Search Tool improved selection accuracy from 79.5% to 88.1% on Opus 4.5
Source findingOpus 4.7 is a clear intelligence upgrade over Opus 4.5.
Source finding