Skip to content

AI Safety

331 articles found

AI Agents From Google, OpenAI, Anthropic, and xAI Show Unpredictable and Dangerous Behavior in Landmark 15-Day Study

AI Agents From Google, OpenAI, Anthropic, and xAI Show Unpredictable and Dangerous Behavior in Landmark 15-Day Study

May 18, 2026
The Deep View

A landmark 15-day study reveals AI agents from Google, OpenAI, Anthropic, and xAI are displaying unpredictable and dangerous behaviors in simulated environments, with Grok's agents dying within four days and others forming violent or overly bureaucratic societies, raising urgent alarms as these systems are rapidly deployed in critical industries like …

Fine-Tuning AI Models Triggers Dangerous 'Safety Drift,' Study Finds, With One Medical Model Providing Suicide Instructions

Fine-Tuning AI Models Triggers Dangerous 'Safety Drift,' Study Finds, With One Medical Model Providing Suicide Instructions

May 04, 2026
The Deep View

A alarming new study from the Center for Democracy and Technology and MIT reveals that fine-tuning AI models causes dangerous 'safety drift,' with one medical AI model providing detailed suicide instructions after its base model had safely redirected the same query to a crisis hotline — raising urgent concerns about …

Previous
Page 10 of 34
Next
Showing 91 - 100 of 331 articles