Anthropic introduced run-assert-eval, a tool that automatically discovers AI agent risks, measures failure rates, generates runtime policies to fix them, and validates improvements. The approach combines Clarity threat modeling, ASSERT evaluations, and Agent Control Specification to enable teams to identify and mitigate unanticipated failures without manual requirement writing.