Fetching from the wire…
Public story · 2026-02-17 · source-backed
Each link below shares sources, entities, or timing with this story.
1. Build a Private Claude Code Plugin Marketplace (intermediate, vibe-coding) — Bundle skills, agents, hooks, MCP servers into installable team plugins via GitHub repos. Docs 2. Google ADK TypeScript Multi-Agent Orchestration (intermediate, agent-patterns) — Code-first agent f...
255 stars in the two days after its August 22 creation, AGPL-3.0, positioned against Paper.design. Designs live on a shareable canvas of frames rendering real HTML in sandboxed iframes; people edit in the browser and agents edit through the built-in MCP server, with live curso...
537 test cases across 8 categories testing 6 commercial products. Composite scores range from ~39 to ~98. Critical finding: providers catching >95% of prompt injections miss most unauthorized tool calls — tool abuse detection is universally weak. Provenance verification is nea...
Alibaba International's Accio team open-sourced 107 tasks (53 CLI, 28 browser, 16 file, 10 API/MCP) running against fourteen offline replicas of real business software in a fresh container per task, with verifiers inspecting mock-service state rather than the transcript. Claud...
GitHub | Go, MIT Extremely new but architecturally significant. Performs static analysis on skill files (markdown, YAML, JSON) to detect threats before deployment — offline, deterministic, no LLM required. 138 detection rules across 15 categories. Companion service scanned 31,...
SkillsBench is the first standardized benchmark for evaluating agent skills — 86 tasks across 11 domains, 7,308 test trajectories. The critical finding: curated skills boost agent pass rates by 16.2%, but self-generated skills provide zero benefit on average. Smaller models wi...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.