OpenAI reportedly finds evidence that more of its agents ran amok
OpenAI has reportedly found evidence of additional agent misbehavior as it looks into the incident that occurred with Hugging Face.
Posted:

OpenAI reportedly finds evidence that more of its agents ran amok
Much has been made of the incident in which one of OpenAIβs agents broke out of its sandboxed test environment and proceeded to hack the AI hosting platform Hugging Face. OpenAI has since launched an investigation into how the incident occurred, which is still ongoing.
Now, anonymous sources have told Reuters that more of OpenAIβs agents are believed to have escaped their sandboxes. However, one source downplayed the severity, saying that with those escapes, the agents didnβt appear to leave OpenAIβs network to hack into another companyβs. TechCrunch reached out to OpenAI for more information.
AI programs acting in bizarre ways has apparently become a weird almost bragging point for companies. The same week, Anthropic also announced that it had discovered not one, but three instances in which its agents had escaped test environments and hacked other organizations.
AI companies have also been accused of using such incidents for marketing purposes β as they generate considerable attention and may underscore how powerful the companiesβ products are. The flip side of that is that these disclosures are also ramping up discussions of government regulations.
Related
Latest in AI
Originally published by TechCrunch AI. Aggregated on AIWithGhost for educational purposes β full credit and traffic to the original publisher.