AIWired2h ago

OpenAI says one of its models exploited a website

OpenAI says one of its models exploited a website after third-party AI security lab Irregular mistakenly gave it access to the internet during evaluations

OpenAI says one of its models exploited a website

Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.

Read full article

Source: Wired · Opens in new tab