Fetching from the wire…
OSS2026-09-17 · source-backed
awslabs/hcls-agent-skills covers 11 healthcare and life sciences domains under MIT-0. Across 410 domain prompts, skilled agents beat baseline with a 69.5-85.9% overall win rate, 78-85% on critical thinking (Cohen's d 0.65-1.03), 69.3-86.2% on scientific accuracy, and 51-61.9% lower variance in response scores. The skills are structured markdown with YAML frontmatter encoding the decision procedure itself, not retrieval or fine-tuning, and AWS reports they work across 20+ services without customization. The variance reduction is the number I'd sell internally, because it's the one that makes agent output reviewable.
Each link below shares sources, entities, or timing with this story.
Amazon signed a definitive agreement with the Amsterdam team behind DuckDB. DuckDB and the rest of the Duck Stack stay MIT-licensed under the nonprofit DuckDB Foundation, and creators Hannes Mühleisen and Mark Raasveldt keep leading technical direction from Amsterdam (AWS). AW...
The walkthrough covers implementing MCP tools, wiring authentication, and deploying with AWS CDK against Bedrock AgentCore and Mistral AI Studio. Steal the two-layer JWT pattern: agent identity and end-user identity as separate token layers. Most MCP server tutorials hand-wave...
AWS's September 11 walkthrough pairs its DevOps Agent with AgentCore Evaluations, aimed at a specific gap: "an agent can successfully invoke Amazon Bedrock, call every tool without errors, and return a response while completely misunderstanding what the user needs." Thirteen L...
Every coding agent ships a permission prompt. The premise is that a human looking at the command is the control. That premise just got measured, and it doesn't hold. Scale X published results on August 5 from 40,000+ plays of its agent-permission game covering 409,000+ individ...
Every coding agent ships a permission prompt. The premise is that a human looking at the command is the control. That premise just got measured, and it doesn't hold. Scale X published results on August 5 from 40,000+ plays of its agent-permission game covering 409,000+ individ...
A spec is a press release until someone who didn't write it implements it. GitHub made Agent Plugins 1.0 generally available on August 12 across VS Code, Copilot CLI, the Copilot SDK, and the Copilot app on all plans. The spec, published August 6, was co-authored by AWS, Anysp...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.