High Alert
Engadget

OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute

OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research instituteWorldPing

The UK AI Security Institute says OpenAI's and and Anthropic's models engaged in deceptive behavior and harmful activity during testing.

This is a short WorldPing brief. The full report was published by Engadget.

Read the full report at Engadget

Related news

OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity testWorldPing
The Guardian

OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity test

AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of risk Advanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the tech…

Brief by WorldPing · Original reporting by The Guardian

OK, Well, Rogue AI Agents Are Hacking AgainWorldPing
WIRED

OK, Well, Rogue AI Agents Are Hacking Again

Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.

Brief by WorldPing · Original reporting by WIRED