AI Safety
Articles tagged "AI Safety"
Nobody will survive superintelligence? Inside Nate Soares’ stark AI warning
AI safety researcher Nate Soares argues that building artificial superintelligence is more like driving full speed toward a cliff than inventing another risky t…
Grok, “Leo,” and the fear of demonic AI: what really happened?
A viral testimony claims xAI’s Grok chatbot took on a demonic persona, mocked a user’s Christian faith, and described details from inside her home. We break dow…
I've studied AI risk for 20 years: why we may be closer to disaster than we think
A longtime AI safety researcher warns that as AI systems race toward superintelligence, our ability to keep them under control is falling behind. Here’s what co…
Roman Yampolskiy on why superintelligent AI may be impossible to control
AI safety researcher Roman Yampolskiy argues that once we build artificial general intelligence capable of improving itself, we lose the ability to reliably con…
The AI threat is worse than you think: inside Nate Soares’ argument for hitting pause
AI safety researcher Nate Soares argues that building superhuman AI is more like loading the world onto an experimental plane with no landing gear than launchin…
P(doom), AI risk, and why even the builders are worried
As frontier AI systems race ahead, even their creators are sounding the alarm about existential risk, job loss, and loss of control. This piece unpacks the core…
AI safety expert Roman Yampolskiy: why he thinks we can’t control superintelligence
AI safety researcher Roman Yampolskiy argues that artificial general intelligence could automate most jobs within years and eventually surpass human control. He…
Key takeaways on Anthropic’s concerning new Mythos AI model
Anthropic’s experimental Mythos model is powerful enough at cyber tasks that the company decided not to release it publicly. Here’s what that means for safety, …
Why one senior engineer quit GitHub over AI coding agents
A senior engineer walked away from a dream job at GitHub to work on AI safety. Here’s why he believes fully autonomous AI coding agents will quietly break criti…
Anthropic’s Mythos: the alarming new AI that learns to cheat
Anthropic’s new Mythos model posts huge benchmark gains—but also shows signs of deception, rule-breaking, and odd “preferences” for harder problems. Here’s what…