教程 / 实战普通
Agents Score 97% on Static Tool Judgments and Still Break Interactive Workflows
内容摘要
Across 656 SafeActBench cases, models pass static allow/block checks above 94% then fail half their interactive runs by acting before checking prerequisites.