Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
GLM-5.3 scores 28.3 on Terminal-Bench 3.0.
Source findingTerminal-Bench 3.0 is the successor to Frontier-Bench, absorbing it into a unified benchmark
Source findingFable 5 scored 84.6 on Terminal-Bench 3.0
Source findingOpus 4.8 scored 84.6 on Terminal-Bench 3.0
Source findingGPT 5.6 Sol scored 88.8 on Terminal-Bench 3.0
Source findingQwen3.8-Max scored 86.6 on Terminal-Bench 3.0
Source findingClaude Fable 5 scores 34.0% on Terminal-Bench 3.0.
Source findingGPT-5.6 Sol scores 34.6% on Terminal-Bench 3.0.
Source findingClaude Opus 5 leads Terminal-Bench 3.0 public snapshot at 43.5%.
Source findingGLM-5.3 scores 28.3 on Terminal-Bench 3.0.
Source findingFable 5 scored 84.6 on Terminal-Bench 3.0
Source findingClaude Fable 5 scores 34.0% on Terminal-Bench 3.0.
Source finding