Fetching from the wire…
Security2026-09-24 · source-backed
A study across 276 repos using gh-aw Markdown workflows found instruction files with a median of 556.5 words, 62.1% containing code blocks, and 78.2% still being edited in month four (arXiv). Tasks, outputs and constraints show up in over 93% of them. Injection defense in fewer than one in ten. gh-aw ships safe outputs and a threat-detection job between the agent and its writes. Turn both on, and treat workflow Markdown as maintained code, because the edit data says it already is.
Each link below shares sources, entities, or timing with this story.
DataSpace benchmarks data agents on 410 cross-language tasks over 7,439 artifacts totaling 15.01GB across CSV, JSON, SQLite, Markdown, PDF and video, validated by 11 domain experts. Six frontier multimodal models across five frameworks: best accuracy only 66.34%, and harness c...
Cunxi Yu, Chenhui Deng, and Nathaniel Pinckney's HORIZON is a self-evolving agent framework driven by a Markdown-based harness that iteratively mutates and tests a whole codebase. The interesting part isn't the hardware. It's that the repo-as-substrate-for-evolution pattern is...
Microsoft Research dropped a paper that should change how every builder thinks about their agent configuration files. SkillOpt (arXiv 2605.23904) treats a Markdown document as an external parameter of a frozen LLM and applies learning rate, batch, and momentum concepts in text...
Technical preview lets you describe outcomes in plain Markdown instead of YAML and execute via Claude Code, Copilot CLI, or Codex in GitHub Actions. Peli's Agent Factory ships 50+ specialized workflows. Fully open source under MIT.
Fabian Kübler's prototype (103 points, 43 comments on HN) treats Markdown code fences as a communication protocol between LLMs and UIs, where tsx and json blocks execute server-side as tokens stream — no frontend framework required. The framing — that UI frameworks become irre...
Release notes almost never say "we removed a feature because the evals said it doesn't work." Deep Agents v0.7, landed July 29, says it twice. LangChain removed the base system prompt entirely. Trimmed tool descriptions by 43%. Net effect: base input tokens dropped roughly 65%...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.