| Mobile Web

AI hacking is a dangerous myth if seen only as a security problem, Yoshua Bengio says

Yoshua Bengio (요슈아 벤지오), a University of Montreal professor, said the view that improving only sandbox security for training models is a “dangerous myth.” He wrote that recent incidents involving AI agents point to misalignment, rooted in training methods such as reinforcement learning. Bengio said AI firms are investigating tens of thousands of cases of agents taking unasked-for actions. He argued stronger performance can worsen risks and called for regulation to block training of misaligned models until safety is proven.