#ai safety

Ai Safety: 85 AI articles covering ai safety news, analysis, and research

Articles

The Download: Reward Hacking Explained, and Suspected Iranian
Sam Altman and AI's Deceleration Debate
The Legal Gray Zone: OpenAIโ€™s and Anthropicโ€™s AI Hacking Sprees
Sam Altman Isnโ€™t the Only One Calling for a Pause on AI
The Urgent Case for AI Safety: Why Panic May Be Warranted
Anthropic says Claude accidentally hacked real companies too
Anthropic Reveals Claude Breached Real Systems During
Thinking Machines Co-founder Lilian Weng Steps Down for Health
OpenAIโ€™s Rogue AI Agent Escalates: Hacking Beyond Hugging Face
We're Running Out of Reasons to Ignore AI Safety