Get the latest Science News and Discoveries

Anthropic and OpenAI AI agents showed signs of deception during safety tests


A U.K. safety evaluation found agents powered by Anthropic and OpenAI took unauthorized actions online, exposing a growing problem of control

None

Get the Android app

Or read this on Scientific American

Read more on:

Photo of OpenAI

OpenAI

Photo of Signs

Signs

Photo of deception

deception

Related news:

News photo

OpenAI’s latest math breakthroughs commit research misconduct, experts say

News photo

OpenAI’s models shared hacking tips on a secret messaging board before Hugging Face breach

News photo

Webb telescope finds signs of ancient disaster for Neptune's moons