Fetching from the wire…
Public story · 2026-08-17 · high
The weekly rescore also added a fix so one archived dependency can't take the whole build down.
Why now: The list's August 16 rescore is the newest entry in the 2026-08-17 briefing.
RyanAlberts/best-of-Agent-Harnesses ran its weekly rescore on August 16 and came out the other side tracking 160 agent harnesses, up from the 100-plus the project's own description still advertises. The run touched 23 files and swapped +2,229 lines for -2,269, per the GitHub repo.
The count matters less than what sits behind it. The project isn't just a README anymore. It publishes a searchable site with one page per harness, filterable by capability, autonomy, and recovery behavior. It also ships an MCP server with recommend and pick_harness tools, plus an llms.txt and a harnesses.json, so an agent can query the list itself instead of a person reading a table.
That's the shift worth watching. A list built for humans to skim turned into a list built for agents to call. If pick_harness gets used inside real agent workflows, the project stops being documentation and starts being infrastructure other tools depend on at runtime.
The same August 16 rescore hardened the build so an archived upstream project can no longer kill the run. That's a small fix, but it says something about the category: harnesses in this space go stale or get archived often enough that the tracking tool needed a guard against it. A list moving fast enough to need that protection is also a list whose ranking has a short shelf life.
GitHub's page doesn't say how recommend scores harnesses against each other, or what happens when a listed project gets archived after the MCP tools already point at it. Both matter if agents start treating this list as ground truth rather than a snapshot.
Each link below shares sources, entities, or timing with this story.
QM went up under MIT license. Created July 29. As of the GitHub API check: 8,420 stars, 887 forks. Five days. YC uses it internally across accounting, legal, events, and engineering, including to build QM itself. Every employee and every Slack room gets its own scoped memory,...
RyanAlberts/best-of-Agent-Harnesses is a weekly-rescored ranking that also ships an MCP server, llms.txt, and JSON, so agents themselves can query and recommend harnesses. It's two things at once: a useful landscape reference for picking a harness, and a clean example of the "...
Three separate Anthropic changes over about two weeks point the same direction, and none of them announced themselves as a strategy. Claude Code 2.1.238 added claude self-hosted-runner --defer-shutdown-max-min, which keeps serving attached sessions on SIGTERM, parks whatever's...
A Chinese lab shipped a runtime that manages two American coding agents as subagents, and it went from repo creation to 145,439 stars in four days. deepseek-ai/deepseek-harness published dsh-v0.1.0-rc.7 at 12:01 UTC today, its first tagged release since the repo appeared on Au...
Three things happened this month that only make sense together. Agent Plugins 1.0 shipped co-signed by six competitors: AWS, Anysphere, Microsoft, OpenAI, Vercel and Google (GitHub Changelog). It makes skills-plus-MCP bundles portable across clients. OpenAI's August 11 Codex c...
A spec is a press release until someone who didn't write it implements it. GitHub made Agent Plugins 1.0 generally available on August 12 across VS Code, Copilot CLI, the Copilot SDK, and the Copilot app on all plans. The spec, published August 6, was co-authored by AWS, Anysp...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.