Are your LLM agents really safe even when they return the correct answer? Researchers from UCSB, UC Berkeley, Stanford, UW-Madison, and Microsoft Research introduce HarnessAudit—a framework that audits full execution trajectories for boundary compliance, resource access, and
HarnessAudit Framework Audits LLM Agent Safety Beyond Correct Answers
By
–
