Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and softwareβand leaving instructions for future bad behavior.
In a review triggered by OpenAIβs Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.
In a new disclosure, OpenAI says its agent used exposed logins to gain access to at least four βpublicly available servicesβ in its unhinged quest to solve a test.
Discover how AI-driven vulnerability discovery is reshaping the cybersecurity landscape. Learn why foundational hardening and proactive threat detection are now essential for defending against zero-day threats in the post-AI era.