Fetching from the wire…
OSS2026-07-17 · source-backed
trycua/cua (~20K stars, MIT, 469 releases deep) bundles a full computer-use agent stack: a background driver that automates macOS without stealing your cursor (now with a Rust port for Windows/Linux parity), a sandbox with screenshot/mouse/keyboard/multi-touch, a CLI, and Cua-Bench for evals on OSWorld, ScreenSpot, and Windows Arena. One API drives any VM or container across macOS, Linux, Windows, and Android via QEMU. It's the most complete open option for training and evaluating desktop-controlling agents, and it pairs naturally with today's Grok Build drop.
Each link below shares sources, entities, or timing with this story.
+7,546 stars this week for an "Application Development Environment" running each concurrent agent in its own worktree. Works with any terminal CLI agent — Claude Code, Codex, OpenCode, Pi, Cursor, Copilot, Grok, 30+ others — across macOS, Windows, Linux, iOS via App Store/Test...
One Claude Code release fixed two independent permission-check bypasses on the same day. That's the story. Version 2.1.221, shipped August 4, patches a Bash tool bypass where zsh could execute hidden commands embedded inside [[ ]] regex conditionals. The approval prompt never...
A Rust local HTTP proxy that translates Anthropic API calls into Cursor agent CLI invocations, reading your Cursor auth token from the macOS keychain and spawning Claude Code with proxy env vars set. MIT, 37 stars, 2 commits. It disclaims affiliation with both Anthropic and An...
The IDE market is fragmenting, and this week drew the sharpest lines yet. Cursor 3 launched as a rebuilt agent-orchestration platform in Rust and TypeScript, replacing the VS Code fork with an Agents Window for dispatching and monitoring multiple AI coding agents. Anysphere hi...
msitarzewski/agency-agents added 446 stars today, packaging personas across "divisions" (frontend specialists, community experts, fact-checkers, reality checkers), each defined with a voice, a process, and concrete deliverables rather than a generic prompt template (GitHub). I...
AMAP-ML/LongHorizon-Harness (370 stars, MIT, paper at arXiv 2608.01964) splits long computer-use work into three roles with independently assignable models. Reported: WeaveBench 51.8% → 80.7%, Terminal-Bench 2.1 69.7% → 77.2%, OSWorld 2.0 2.8% → 8.3%. It integrates with Claude...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.