Wire
RegulatoryAgentBench tests agents across 50 rules
RegulatoryAgentBench evaluates AI agents on 50 compliance scenarios drawn from more than 30 regulators across APAC, EMEA, LATAM, and North America. The open benchmark repository scores action coverage, fact extraction, deadline accuracy, and critical misses, while noting that its scenario-generation pipeline depends on a private annotation dataset. Compliance teams can use the harness to test the full read-to-action loop, but should not mistake a 50-scenario sample for jurisdiction-wide assurance; the archive’s agent-evaluation governance analysis supplies the stricter acceptance boundary.