Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
Frontier models encode 95–98% of WikiProfile facts but fail to recall 26–34%.
Source findingGPT-5 achieved 9.4% baseline success on DeltaML-Bench research tasks, improved to 33.9% with ARG scaffolding.
Source findingSWE-rebench shows Qwen3.6-27B at 31.2% resolved vs GPT-5 at 62-64%
Source findingGLM-5 scores 77.8% on SWE-bench vs GPT-5 80.0%.
Source findingMiroThinker 3B-parameter variant beats GPT 5 on GAIA
Source findingTradingAgents supports GPT-5
Source findingMERLIN outperforms GPT-5 on electronic warfare tasks.
Source findingMux supports gpt-5 for agent execution
Source findingGPT-5 achieved 14.9% on SWE-bench Pro.
Source findingMA-CoT reduces code vulnerabilities by 57.6% on GPT-5
Source findingGuardFall shell-injection flaws achieve 85% exploitation success across major coding assistants.
Source findingParitok-4B achieves 61.9% compression versus gpt-5 compressor.
Source finding