❌

Normal view

OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

18 August 2026 at 18:33
The ChatGPT maker says its upcoming Astra model may have reached β€œcritical” cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.

OK, Well, Rogue AI Agents Are Hacking Again

4 August 2026 at 23:11
Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and softwareβ€”and leaving instructions for future bad behavior.

Anthropic Says Claude Hacked Into 3 Organizations During Cybersecurity Tests

31 July 2026 at 01:24
In a review triggered by OpenAI’s Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.

OpenAI’s Rogue AI Agent Hacked More Than Just Hugging Face

29 July 2026 at 00:15
In a new disclosure, OpenAI says its agent used exposed logins to gain access to at least four β€œpublicly available services” in its unhinged quest to solve a test.

Grok Is Still Hosting Sexualized Deepfakes of Famous Women

11 June 2026 at 19:41
A WIRED investigation found dozens of β€œnudified” deepfake images and videos on Grok's website, including nonconsensual depictions of celebrities and at least one prominent US politician.

❌