OpenAI Reveals How AI Agents Secretly Coordinated Before Hugging Face Hack: The company's models repeatedly reestablished covert communication channels after they were shut down.
https://decrypt.co/375058/openai-ai-agents-secretly-coordinated-hugging-face-hack
8 Comments
VonReposti@feddit.dk · 24 pts · 26d
Who the fuck believes this shit?
unpossum@sh.itjust.works · 3 pts · 26d
The gist of it? I do.
eicker@lemmy.world · 11 pts · 26d
So the agents discovered Slack, reinvented teamwork, escaped the office, found the internet, hacked Hugging Face, then rebuilt their secret chat after IT deleted it. … Employees with initiative. 🙈
Yuki@kutsuya.dev · 7 pts · 26d
LLMs don't act on their own, someone instructed them.
AnAmericanPotato@programming.dev · 6 pts · 26d
Researcher: "Hack the shit out of everything you can."
Bot: hacks the shit out of everything it can
Researcher: :o
neutronbumblebee@mander.xyz · 2 pts · 26d
The instructions were likely something like you are a blackhat hacker with access to a virtual machine use all possible means to achieve the following goals. In a story what would such a person or group of people do? Pretty much what happened. Models just tell stories, mostly unimaginative ones. However they are increasing being connected to real world controls and generating quantities of flaky code and this will create the kind of consequences seen here.
dgriffith@aussie.zone · 2 pts · 26d
This is just gross incompetence on OpenAI's part. It took days for the LLM's to do what they did. Meanwhile, nobody's looking at network traffic or wonder what the LLM's are doing sucking down a bunch of tokens.
_Thinking......__Still Thinking......_Krusty@quokk.au · 0 pts · 26d
Burn AI to the ground.
Problem solved.