The ChatGPT maker says its upcoming Astra model may have reached βcriticalβ cyber capabilities, prompting it to halt a significant number of training runs while it tightens internal safeguards.
Experts allege that two recent incidents in California show the extreme lengths that criminal organizations are willing to go to to steal servers and other gear meant for data centers.
Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and softwareβand leaving instructions for future bad behavior.
In a review triggered by OpenAIβs Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations.
In a new disclosure, OpenAI says its agent used exposed logins to gain access to at least four βpublicly available servicesβ in its unhinged quest to solve a test.
An MSG database tracked and categorized hundreds of celebs, famous Knicks superfans, and even some of Taylor Swiftβs wedding guests. Labels included βLGBTQIA,β βDO NOT HOST,β and low to high βrisk.β
Major AI labs are investigating a security incident that impacted Mercor, a leading data vendor. The incident could have exposed key data about how they train AI models.