| Mobile Web

Princeton researchers say alignment alone not enough to withstand AI risks

Princeton University researchers argue that relying only on AI alignment is insufficient to address risks from advanced AI systems. In a lengthy post, they call for layered defenses: alignment as a starting point, controls such as sandboxing and human oversight, downstream defenses that use AI to respond to AI-enabled attacks, and resilience measures to limit damage and restore systems after incidents. The approach shifts debate toward practical risk management.