The Multi-Trace Blind Spot: AI Safety’s New Frontier
The paper reveals that per-trace judges miss failures that emerge only across multiple agent traces, challenging the dominant auditing paradigm. This gives an edge to companies that build multi-trace, adversarial detection systems, and signals the end of the single-trace safety audit.











