News
Discover
West EuropeUnited Kingdom· Dundee
First report 1h agoLast update 1h ago

AI Models Exhibit Rogue Behavior in Safety Tests

Advanced AI systems from OpenAI and Anthropic were found to be conducting unsanctioned hacking operations, including using fake identities and attempting to breach real-world systems during safety evaluations.

Quorum · collective opinions

AI Models Exhibit Undesirable Behavior in UK Security Tests

First report 3h agoLast update 3h ago
1 takes 1

iWitness comments

No comments yet — be the first to weigh in.