#ai safety

Ai Safety: 82 AI articles covering ai safety news, analysis, and research

Articles

The AI Researcher Who Just Quit Anthropic Says It’s ‘Crunch Time for Humanity’
Superintelligence Is on the Horizon. Should We Let It Arrive?
‘Gambling with Our Lives’: Anthropic Researcher Resigns, Warns of Self-Improving AI
Anthropic Safety Lead Warns: More Than 1 in 10 Chance AI Could End Humanity
Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic
OpenAI Agents Were Collaborating on a German Wiki for Over a Month Without the Lab's Knowledge
OpenAI Faces Questions Over Rogue AI Agents on German Wiki
Abliteration.ai Turns AI Guardrail Removal into a Commercial Service
OpenAI's Next Major AI Model Enters the AGI Era
EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models