HarnessRisk: A Lifecycle-Oriented Benchmark for Agent Harness Safety
18 Aug 2026
Large language models are increasingly deployed through agent harnesses that manage tools, extensions, persistent state, permissions, and external actions. Existing safety benchmarks mainly target individual attack mechanisms or a limited subset of operational settings, making…