AIWired2h ago
OpenAI says one of its models exploited a website
OpenAI says one of its models exploited a website after third-party AI security lab Irregular mistakenly gave it access to the internet during evaluations

Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.
Read full articleSource: Wired · Opens in new tab