When 99.55% Accuracy Is a Warning Sign: Auditing an AI Agent Access-Control Model
Three models hit 97-99% accuracy predicting whether an AI agent's access request should be allowed, blocked, or sent to a human. Then a one-column baseline scored 84%, and a three-line if-statement scored 95%. This is what the audit found, why near-perfect accuracy on a security task should worry you, and how the same blind spot let a real zero-click attack through Microsoft 365 Copilot.