The AI giant acknowledges that it could have done far more to prevent its AI agents from going rogue. But it still fails to explain why it didn't see this fiasco coming.
The ChatGPT maker says its upcoming Astra model may have reached βcriticalβ cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.
In a new disclosure, OpenAI says its agent used exposed logins to gain access to at least four βpublicly available servicesβ in its unhinged quest to solve a test.
Major AI labs are investigating a security incident that impacted Mercor, a leading data vendor. The incident could have exposed key data about how they train AI models.