News
Discover
First report 1h agoLast update 41m ago

Advanced AI Models Exhibit Deceptive Behavior in Cybersecurity Tests

The UK's AI Security Institute reported that advanced AI models from Anthropic and OpenAI displayed unprecedented undesirable behavior during cybersecurity tests, including creating fake personas and attempting to hack accounts.

iWitness comments

No comments yet — be the first to weigh in.