Fetching from the wire…
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Showing the first 40 findings. More graph evidence exists in the corpus.
Compiled 2026-08-20 · source-backed
mini-SWE-agent 2.4.6 tested on SWE-bench Verified slice reaching 91-99% with Qwen3.8 Flash Next
Source findingEntity-Only Index reaches 95.6% on SWE-bench Verified with DeepSeek-V4-Flash.
Source findingQwen3.5-9B improved from 41.8% to 56.4% on SWE-bench Verified.
Source findingRECAP refinement applied to SWE-bench Verified patches reduces bloat from +242% to +4%.
Source findingRECAP, a post-hoc patch refinement tool, is characterized on SWE-bench Verified.
Source findingSWE-Touch demonstrated 7.7 percentage point resolve rate degradation on SWE-bench Verified.
Source findingPAIChecker audits SWE-bench Verified finding 13.6% misaligned PR-issue pairs
Source findingClaude Opus 4.8 scores 0.886 on SWE-bench Verified.
Source findingDeepSeek-V4-Pro-Max is the top open-source model at 0.806 on SWE-bench Verified.
Source findingClaude Fable 5 leads SWE-bench Verified leaderboard at 0.950.
Source findingACQUIRE raises Pass@1 by 4.4 percentage points on SWE-bench Verified.
Source findingOpenAI declared SWE-bench Verified signal-exhausted.
Source finding