skip to content
The Weighted Average

Wire

Niyam-AI makes agent guardrails verifiable

Niyam-AI reports 88.5% F1 with a 1.1% false-positive rate on 2,000 Agent-SafetyBench scenarios, while verifying approved tool calls in 53.1 milliseconds after proof generation. The preprint binds permitted tools and constraints to a SHA-256 intent contract, then requires a zk-SNARK proof before execution; proof generation adds about 2.26 seconds, and the authors caution that the comparison uses a benchmark-adapted classifier against zero-shot baselines. For teams moving beyond prompt-level agent permissions, the operational trade is explicit: verifiable policy can fit a fast control path, but approval latency and benchmark transfer still need a production pilot.