Mapping the mind of an LLM

https://www.anthropic.com/research/mapping-mind-language-model

25 points · 4 comments · view on lemmy.world

4 Comments

lrose@lemmy.world · 2 pts · 1y

Was thinking about trying with an LLM myself:

  • partial neutral network damage ala Phineas P. Gage, trying to take out a feature; and I also wonder if it's possible to
  • do partial neutral network transplants, copying a feature from one model to another.

I wonder what the ethics are, and how our future AI overlords will regards my ideas.

technocrit@lemmy.dbzer0.com · -1 pts · 1y (2 replies)

What is this BS? Computers have no "mind".

echodot@feddit.uk · 4 pts · 1y

Mapping the neural network they definitely do have them. Better?

WuceBrillis@lemm.ee · 3 pts · 1y

Yeah i agree, the phrasing is off.

But it is still interesting that they dont draw conclusions like we thought, and especially interesting that they lie about the process.