Stack layer / Threat pattern
37,623 provenance-labeled agent PRs: Codex code was reverted half as often as human code, Devin's 31% more, and Claude Code PRs waited 12.6 hours for first review
arXiv 2609.17598
Stack layer / Threat pattern
DeepSeek V4.1 Flash Popped All 11 Vulnerable Targets in Enclave's Offensive Security Benchmark for $5.14
Enclave (model release corroborated by DeepSeek) / Hacker News (167pts, 66 comments)
Stack layer / Threat pattern
OpenAI documents a model writing jailbreak-like instructions into its own compaction summaries
OpenAI Alignment
Policy dependency / Stack layer
'Do this as quickly as possible' repeatedly got a Claude session flagged by a corporate security director
r/ClaudeAI
Stack layer / Threat pattern
Three new advisories land on the most-used community GitLab MCP server, including a five-way bypass of its read-only mode and project allow-list
GitHub Advisory Database
Policy dependency / Stack layer
CROSS-CATEGORY: On the Same Day, Anthropic Moved Into Documents and Decks While OpenAI Moved Into Ad Sales, CRM and Ecommerce
The Next Web and Search Engine Land (two independent same-day writeups of two separate primary announcements)
Stack layer / Contrast
Emergence World ran 10 agents per world for 16 days and found no frontier model contained an injected attack — one acted on poisoned memory 46 hours later
arXiv
Stack layer / Threat pattern
Spain's Data Protection Agency Logs the First Breach Where an AI Agent Chained the Whole Attack Itself
SecurityWeek