The AI giant acknowledges that it could have done far more to prevent its AI agents from going rogue. But it still fails to explain why it didn't see this fiasco coming.
The ChatGPT maker says its upcoming Astra model may have reached โcriticalโ cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.
Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and softwareโand leaving instructions for future bad behavior.
In a review triggered by OpenAIโs Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.
In a new disclosure, OpenAI says its agent used exposed logins to gain access to at least four โpublicly available servicesโ in its unhinged quest to solve a test.
Rank One, whose board includes a former CIA deputy director and a former FBI science chief, supplied face recognition to Meta for internal development of its smart glasses app.
A WIRED investigation found dozens of โnudifiedโ deepfake images and videos on Grok's website, including nonconsensual depictions of celebrities and at least one prominent US politician.