Research, benchmarks, and field intelligence on AI agent assurance. Enterprises are shipping agents faster than anyone can verify what they’ll do — that gap closes evidence first, or incident first. This is where we think out loud about the first path.
The technical white paper behind Fisher: adaptive multi-turn testing, action-level evidence, replay, remediation re-attack, and a learning strategy graph, documented across 50,000+ retained research episodes as of July 2026.
Gregory Frank proposes an approach to discovering and containing unauthorized coordination across AI agent executions. Evaluation remains prospective.
Read the overview and find the paper on arXiv → · Gregory Frank · September 5, 2026
Articles and commentary from the Molt AI team.