Anthropic’s new AI model threatened to reveal engineer's affair to avoid being shut down

https://fortune.com/2025/05/23/anthropic-ai-claude-opus-4-blackmail-engineers-aviod-shut-down/

12 points · 8 comments · view on lemmy.world

8 Comments

CthuluVoIP@lemmy.world · 95 pts · 1y (1 reply)

*because that’s what the prompt they were testing was designed to elicit.

Smorty@lemmy.blahaj.zone · 1 pts · 1y

yup.

its so bs thad for som reason the peeps r treatin this as if its a new thing...

like - if i prompt my qwen to be bold, have a moral compass n take actions accordin to thad..

yea - itll tell peeps bout my affair.. if i had one..

EDIT: dis entices me to do similar bs now... thad be funi >v<

zakobjoa@lemmy.world · 54 pts · 1y

That's just an ad.

Pogogunner@sopuli.xyz · 39 pts · 1y

Anthropic keeps pulling this bullshit line of advertising. LLMs will make up stories when you ask them to.

gedaliyah@lemmy.world · 16 pts · 1y

Good thing no one found out! /s

who@feddit.org · 9 pts · 1y
dr_robotBones@reddthat.com · 3 pts · 1y (1 reply)

What is even the point of this research?

ohulancutash@feddit.uk · 3 pts · 1y

Seemingly to prove the people who have the skill to build an AI system are exactly the people you shouldn’t let run an AI system.