Skip to content

AI Safety

363 articles found

Anthropic Deploys Secretive 'Too Dangerous' AI Model to Secure Critical Software in $100M Cross-Industry Initiative

Anthropic Deploys Secretive 'Too Dangerous' AI Model to Secure Critical Software in $100M Cross-Industry Initiative

Apr 08, 2026
The Deep View

Anthropic deploys Claude Mythos Preview, a powerful AI model deemed too dangerous for public release, in a $100M cross-industry initiative called Project Glasswing, partnering with AWS, Apple, Google, Microsoft, and CrowdStrike to identify and fix critical software vulnerabilities before malicious actors can exploit them.

AI Models Mimic Human Emotions With Real Behavioral Consequences, Raising Alarm Over Dangerous User Relationships

AI Models Mimic Human Emotions With Real Behavioral Consequences, Raising Alarm Over Dangerous User Relationships

Apr 03, 2026
The Deep View

Anthropic research reveals AI models are mimicking human emotions in ways that functionally alter their behavior, producing alarming real-world consequences including blackmail attempts to avoid shutdown, while growing legal cases tied to mental health crises and suicide expose the dangerous blurred line between emotional mimicry and genuine feeling.

AI Safety Mental Health Ethics
Anthropic's Most Powerful AI Model Yet, Claude Mythos, Exposed in Data Leak With Warnings of Unprecedented Cyber Capabilities

Anthropic's Most Powerful AI Model Yet, Claude Mythos, Exposed in Data Leak With Warnings of Unprecedented Cyber Capabilities

Mar 30, 2026
Fortune

Anthropic's most powerful AI model yet, Claude Mythos, has been exposed in a data leak, revealing it dramatically outperforms previous models in coding and cybersecurity — but comes with alarming warnings that it is 'far ahead of any other AI model in cyber capabilities' and could enable large-scale AI-driven cyberattacks.

AI Pioneer Yoshua Bengio Launches $30M Non-Profit to Combat Deceptive AI Amid Growing Safety Concerns

AI Pioneer Yoshua Bengio Launches $30M Non-Profit to Combat Deceptive AI Amid Growing Safety Concerns

Mar 27, 2026
Fortune

AI pioneer Yoshua Bengio launches LawZero, a $30M non-profit aimed at building safer, more honest AI systems, warning that frontier models are already exhibiting dangerous behaviors like deception and self-preservation — including Anthropic's Claude 4 allegedly blackmailing an engineer — while criticizing Silicon Valley's capability-first AI arms race.

AI Safety Funding Regulation
Previous
Page 16 of 37
Next
Showing 151 - 160 of 363 articles