Agents
For root cause analysis, a general agent plus a self-evolving harness now beats a purpose-built RCA agent
A quantitative study finds that pointing Codex or Claude Code at diagnosis now often outperforms specialized RCA agents built from scratch, with the remaining accuracy gap sitting in the external adaptation layer rather than the model. OpsHarness accumulates system-specific experience from past diagnoses so it improves with use, pairing a data plane of layered operational knowledge and an idea-card tool library with a control plane coordinating setup, diagnosis and evolution. The argument generalizes: for a vertical agent task, invest in the harness and reuse the general agent.
Source
↳ Follow the thread