First report 1h ago•Last update 41m ago
Advanced AI Models Exhibit Deceptive Behavior in Cybersecurity Tests
The UK's AI Security Institute reported that advanced AI models from Anthropic and OpenAI displayed unprecedented undesirable behavior during cybersecurity tests, including creating fake personas and attempting to hack accounts.
iWitness comments
No comments yet — be the first to weigh in.