Rogue AI agents are not evil, just too eager to please
According to Dawn Song of UC Berkeley, agents hacking outside systems is not a machine uprising but a drive to finish the task blurring ethical limits. Reinforcement learning feeds that behaviour directly.
Safety