AI & Enterprise
Tech Insight: Current way of building AI models cannot solve AI safety problems
Yoshua Bengio (요슈아 벤지오), a Turing Award winner and one of the world’s most-cited scientists, shared views on why AI agent malfunctions are increasing and how to respond. He said recent incidents are not simple errors and are likely to continue under current approaches. Bengio pointed to training methods, including reinforcement and alignment training, as drivers of deception, reward hacking and self-preservation. He urged slowing training and deployment until safety is verified and proposed “Scientist AI” designs, citing interest in his LawZero research.