Anthropic’s new AI model threatened to reveal engineer's affair to avoid being shut down
https://fortune.com/2025/05/23/anthropic-ai-claude-opus-4-blackmail-engineers-aviod-shut-down/
https://fortune.com/2025/05/23/anthropic-ai-claude-opus-4-blackmail-engineers-aviod-shut-down/
8 Comments
CthuluVoIP@lemmy.world · 95 pts · 1y
*because that’s what the prompt they were testing was designed to elicit.
Smorty@lemmy.blahaj.zone · 1 pts · 1y
yup.
its so bs thad for som reason the peeps r treatin this as if its a new thing...
like - if i prompt my qwen to be bold, have a moral compass n take actions accordin to thad..
yea - itll tell peeps bout my affair..
ifihadone..EDIT: dis entices me to do similar bs now... thad be funi >v<
zakobjoa@lemmy.world · 54 pts · 1y
That's just an ad.
Pogogunner@sopuli.xyz · 39 pts · 1y
Anthropic keeps pulling this bullshit line of advertising. LLMs will make up stories when you ask them to.
gedaliyah@lemmy.world · 16 pts · 1y
Good thing no one found out! /s
who@feddit.org · 9 pts · 1y
https://web.archive.org/web/20250526131412/https://fortune.com/2025/05/23/anthropic-ai-claude-opus-4-blackmail-engineers-aviod-shut-down/
dr_robotBones@reddthat.com · 3 pts · 1y
What is even the point of this research?
ohulancutash@feddit.uk · 3 pts · 1y
Seemingly to prove the people who have the skill to build an AI system are exactly the people you shouldn’t let run an AI system.