Fetching from the wire…
Agents2026-06-20 · source-backed
Marginal Advantage Accumulation fixes contradictory cross-batch feedback in trace distillation by building differential signals, accumulating per-operation evidence via EMA, and merging semantic identities for traceability (arXiv:2606.20475). For anyone building agents that learn from their own run history, this is the rare paper that improves quality and slashes cost at once. Worth a close read if your agent's memory layer is getting expensive.
Each link below shares sources, entities, or timing with this story.
MemSyco-Bench points out that memory benchmarks test whether memories are correctly stored, retrieved, and updated, never whether the retrieved memory should have influenced the decision at all. Its five tasks check whether agents can reject memory as factual evidence, respect...
In a GitHub Copilot SDK setup, an asynchronous memory-curator agent got read-only tools to check candidate memories against the current state before saving them. Pass rate on CLBench rose to 73% from 39% (arXiv 2609.11060). Queries per question fell to 4.7 from 8.8, and task-a...
$3,054 against $38,370. Same benchmark, better score. Praxist (arXiv 2608.25955, submitted August 26) replaces per-attempt agent memory with a typed evidence graph of findings, plus lane-structured frontiers and agendas, so later attempts inherit validated mechanisms rather th...
On a verifiable protein-function characterization task routed across tools, model choice swamped federation topology, RL-versus-LLM harness, and prompt expertise: Opus at roughly 92 to 94%, o4-mini at 40 to 50%. Federation across institutional boundaries cost almost nothing (a...
Recuris (arXiv 2608.24876) keeps a Working Memory tracking current task progress separate from an Experiential Memory of learned skills, so skill selection indexes against what the task needs now rather than the whole history. It improves 35 of 37 model-benchmark pairs, gains...
RGA-Designer trains a reward model scoring both task correctness and structural compactness, then fine-tunes a graph generator against it to design communication topologies. arXiv For fan-out agent teams where inter-agent chatter dominates the bill, topology is a cost lever mo...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.