Agents need a control plane outside their reach
Google DeepMind has published an Artificial Intelligence Control Roadmap for securing internal systems as agents become more capable and operate with greater autonomy. The work considers how organizations can manage systems that may be useful, imperfectly aligned, and able to interact with valuable digital resources.
One principle deserves broad adoption: the mechanisms that observe and constrain an agent should not depend entirely on the agent's own willingness or ability to comply.