Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
Compiled 2026-09-10 · source-backed
GPT-5.5 reached 80% F1 on VEX-Bench.
Source findingMythos completed end-to-end attack chain 3/10 times vs GPT-5.5 2/10 in UK AISI testing
Source findingAn auto-research loop drove GPT-5.5 through 1,500+ optimization submissions achieving 232x speedup on a QR kernel.
Source findingOpenAI released GPT-5.5.
Source findingGPT-5.5 was evaluated on SRE-Bench
Source findingGPT-5.6 Sol at low reasoning outperformed GPT-5.5 at high reasoning.
Source findingFrontis-MA1 exceeds GPT-5.5 on code evolution tasks
Source findingKnowAct-GUIClaw outperformed GPT-5.5 on MobileWorld.
Source findingCodex CLI with GPT-5.5 leads the Terminal-Bench 2.1 public leaderboard at 83.4%.
Source findingGPT-5.5 deployed correct backends 28.6% of the time on the BackendForge benchmark.
Source findingHourglass Reasoning achieves 58% Verilog synthesis accuracy with GPT-5.5.
Source findingClaude Fable beat GPT-5.5 which achieved 4.34x speedup
Source finding