#ai safety

Ai Safety: 85 AI articles covering ai safety news, analysis, and research

Articles

Abliteration.ai Turns AI Guardrail Removal into a Commercial Service
OpenAI's Next Major AI Model Enters the AGI Era
EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models
OpenAI’s Astra Model Uses a New Reasoning Technique That Has AI Safety Experts Worried
Google DeepMind Releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber: One Core Model, Two Access Envelopes
Researchers Warn of Safety Risks Ahead of OpenAI's Astra Launch
OpenAI's Astra Model: A Powerful New Tool for Cybersecurityβ€”and a Potential Threat
OpenAI Delays New Model Development Following Hugging Face Breach
The Download: Engineered Microbes for Crops, and OpenAI's Culture Problem
Human-in-the-Loop Without Killing Throughput