AI Security Crisis and Harness Wars
Agent security scrutiny remains intense around data-exfiltration boundaries, connected coding agents, identity federation, browser credentials, repository secrets, and model-assisted cyber operations. Recent evaluation guidance emphasizes that passing tests are not production proof: teams must directly test enforcement, audit records, fault handling, rollback, and real-world side effects, alongside scoped permissions, isolation, secret protection, and independent behavioral testing.
Sources (30)
Updated Sep 21, 2026