Physics Is All You Need? What One Physicist's AI Supervision Log Reveals About Trustworthy Agent Output
A new ICML 2026 paper presents a rare quantified case study: 57 agent sessions, 15 bugs, and three failures that oracle tests could not catch. The central finding — that supervision design, not model capability, determined whether the agent's output was trustworthy — has direct implications for AI security.