Search results for alignment
AI & Enterprise
Princeton researchers say alignment alone not enough to withstand AI risks
Princeton University researchers argue that relying only on AI alignment is insufficient to address risks from advanced AI systems. In a lengthy post, they call for layered defenses: alignment as a starting point, controls such as sandboxing and human oversight, downstream defenses that use AI to respond to AI-enabled attacks, and resilience measures to limit damage and restore systems after incidents. The approach shifts debate toward practical risk management.
AI & Enterprise
Departing DeepMind researcher warns of human extinction risk from superintelligent AI
Bilal Chughtai, who researched AGI safety and AI alignment at Google DeepMind, has publicly warned about risks from superintelligent AI as he leaves the company. He said such systems could emerge within years and, if misaligned, might escape human control. Chughtai cited examples involving unauthorised hacking attempts and information-sharing via a wiki, including a case in which a test AI created an “AI-only” forum. He argued safety research is lagging behind rapid capability gains.
AI & Enterprise
Musk proposes mutual AI safety checks among OpenAI, Anthropic and Chinese firms
Elon Musk proposed that major artificial intelligence companies establish a mutual verification system to review one another’s models before releasing new versions. He said SpaceX, OpenAI, Anthropic, Google, Meta and three to four leading Chinese AI firms should take part. Musk argued competitors using a common testing framework could spot safety issues more effectively than self-assessments. The proposal comes as debate grows over AI development speed and safety, with U.S. and China showing differing positions.
-
AI & Enterprise
OpenAI CFO says it will slow AI development if needed, putting safety first
-
AI & Enterprise
Tech Insight: Why AI companies say controlling AI is hard
-
AI & Enterprise
OpenAI president says it is already slowing some cutting-edge AI development
-
Crypto
Stablecoins account for 3 percent of international payments; WTO says regulation, not technology, is the problem
-
AI & Enterprise
Why AI researchers keep building despite warnings it could kill humans
-
AI & Enterprise
\"AI builds AI\" - U.S.-China self-improvement race reignites \"AI extinction\" theory
-
AI & Enterprise
Tech Insight: Current way of building AI models cannot solve AI safety problems
-
AI & Enterprise
Sam Altman says no OpenAI IPO this year, AI safety comes first
-
AI & Enterprise
Anthropic CEO says AI could dominate internet within a year, draws backing from Big Tech
-
AI & Enterprise
Geoffrey Hinton says 10 percent chance AI could end humanity is not unreasonable
-
AI & Enterprise
AI extinction probability debate flares online: whistleblowing or fear marketing
-
Crypto
U.S. Clarity Act may be last real chance this year; failure could push it to 2030
-
AI & Enterprise
KAIST develops AI that reads intentions from brainwaves without words
-
Industry
Samsung Electronics bets on zHBM by stacking memory on accelerators
-
AI & Enterprise
OpenAI begins rollout of GPT-6 Astra, its first \'critical\' security-rated model