Research
AgentDV Lifts RTL Testbench Generation From Zero Valid Environments to 100% Pass on Four DUTs
Single-shot LLM prompting failed to produce even one valid coverage-producing verification environment on the paper's benchmarks. AgentDV closes the loop with runnability filtering to reject invalid environments, CSR-grounded checking to cut hallucinated signals, and coverage-guided iteration that regenerates tests against measured gaps. Using Claude Sonnet 4.6 it reaches 100% pass on four design-under-test blocks and 80.9% average across all of them, with 74.5% line and 88.4% branch coverage, against 58.7% and 60.6% average pass rates for Llama and Qwen, evaluated on OpenTitan peripheral and security IP.
↳ Follow the thread