Fetching from the wire…
Tools2026-09-26 · source-backed
The tool, created September 25 as a Claude Code plugin, GitHub Action and CLI, runs each changed pytest, vitest or jest test twice: once with the change, once with only source files reverted. Each test gets labelled PROVEN, THEATER (passes both ways) or WEAK. Across 181 real changes, 82% of agent PRs were proven against 90% of maintainer fixes, and in 10% of agent PRs every test failed on old code only because it imported a new name. Its prove-fix skill makes Claude rewrite the test, not the fix, until it's proven. This is the cleanest answer I've seen to "the agent wrote tests, but do they test anything."
Each link below shares sources, entities, or timing with this story.
1. Set package cooldown to 72 hours across all your package managers. pnpm: resolution-time=72h, uv: --exclude-newer, npm via .npmrc. This single config change would have protected you from the LiteLLM attack. Willison's survey covers all seven managers. 2. Install Lasso Secur...
DietrichGebert/ponytail cut v4.10.0 on September 14 at 138,988 stars, with a scripts/cursor-hooks.js install that merges into an existing Cursor hooks file without clobbering the rest (GitHub). Same release fixes VS Code Copilot detection via a CLAUDE_PLUGIN_ROOT fallback and...
The IDE market is fragmenting, and this week drew the sharpest lines yet. Cursor 3 launched as a rebuilt agent-orchestration platform in Rust and TypeScript, replacing the VS Code fork with an Agents Window for dispatching and monitoring multiple AI coding agents. Anysphere hi...
alibaba/open-code-review cut v1.12.1 at 11:23 UTC this morning. It's a Go CLI that ran as Alibaba's internal review assistant for two years before the May 2026 open-source release, and it's at 24,596 stars with 443 added today. The architecture explains the claimed numbers. It...
bradautomates/claude-video (v0.2.0, July 1) lets agents download, frame-extract, and transcribe any video via yt-dlp, ffmpeg, and Whisper, then hand it to Claude's multimodal Read (GitHub). It ships as an Agent Skill usable across 50+ agents: Claude Code, Codex, Cursor, Gemini...
This one changed how I'm spending my week. Anthropic's July 24 context-engineering post says they removed over 80% of Claude Code's system prompt for Opus 5 and Fable 5 with no measurable loss on coding evals. They call it "unhobbling" — stripping guardrails and rules that new...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.