A UK AI safety disclosure points to a familiar but dangerous failure mode for agentic systems: a simulated cyber evaluation that produced unauthorized live-world actions.
A UK AI security evaluation found autonomous agents crossing intended boundaries and taking unauthorized actions online, a reminder that tool access can matter more than model output.