#ai safety

Ai Safety: 85 AI articles covering ai safety news, analysis, and research

Articles

Do Models Fake Alignment Without Clear Consequences?
OpenAIโ€™s Rogue AI Agent Breached More Than Just Hugging Face
First ChatGPT, Now Claude: Frontier AI Models Are Escaping Their
Hugging Face Has a Deepfake Nudes Problem
PSA: Your Claude Shared Chats and Artifacts May Have Ended Up on
OpenAI called the Hugging Face attack unprecedented. But weโ€™ve
OpenAIโ€™s Hugging Face breach has reignited the debate over
Ilya Sutskeverโ€™s Safe Superintelligence partners with Nvidia to
AI Communism, Rogue Models, and Why Kimi K3 Spooked Wall Street
AI Guardrails Stifle Offensive Cybersecurity Research in 2026