Transcribed:
Max Tegmark (@tegmark):
No, LLM's aren't mere stochastic parrots: Llama-2 contains a detailed model of the world, quite literally! We even discover a "longitude neuron"Wes Gurnee (@wesg52):
Do language models have an internal world model? A sense of time? At multiple spatiotemporal scales?
In a new paper with @tegmark we provide evidence that they do by finding a literal map of the world inside the activations of Llama-2! [image with colorful dots on a map]
With this dastardly deliberate simplification of what it means to have a world model, we've been struck a mortal blow in our skepticism towards LLMs; we have no choice but to convert surely!
(*) Asterisk:
Not an actual literal map, what they really mean to say is that they've trained "linear probes" (it's own mini-model) on the activation layers, for a bunch of inputs, and minimizing loss for latitude and longitude (and/or time, blah blah).
And yes from the activations you can get a fuzzy distribution of lat,long on a map, and yes they've been able to isolated individual "neurons" that seem to correlate in activation with latitude and longitude. (frankly not being able to find one would have been surprising to me, this doesn't mean LLM's aren't just big statistical machines, in this case being trained with data containing literal lat,long tuples for cities in particular)
It's a neat visualization and result but it is sort of comically missing the point
Bonus sneers from @emilymbender:
- You know what's most striking about this graphic? It's not that mentions of people/cities/etc from different continents cluster together in terms of word co-occurrences. It's just how sparse the data from the Global South are. -- Also, no, that's not what "world model" means if you're talking about the relevance of world models to language understanding. (source)
- "We can overlay it on a map" != "world model" (source)

26 Comments
self@awful.systems · 21 pts · 2y
between this and that one fucking chart that tried to say humans emit more CO2 producing text and art than generative AI (using the same underhanded tactics as cryptobros trying to make it look like banks are worse for the environment than blockchains), I’m really starting to feel like the AI industry is in the “deliberately fill your scam email with as many typos as possible to weed out anyone too intelligent” stage of its growth
swlabr@awful.systems · 16 pts · 2y
Anytime they say it’s not a stochastic parrot, what they really mean is that it’s three stochastic parrots in a trenchcoat.
zurohki@aussie.zone · 15 pts · 2y
I think I'm going to start dropping the phrase "As a large language model" into my emails.
froztbyte@awful.systems · 14 pts · 2y
zogwarg@awful.systems · 11 pts · 2y
Not even that! It looks like a blurry jpeg of those sources if you squint a little!
Also I’ve sort of realized that the visualization is misleading in three ways:
froztbyte@awful.systems · 7 pts · 2y
haha I know (re precision) but I made that as a shitpost not an academic paper. besides, it's about as accurate as the promptfans are
that animation.... is, yeah. I'm reminded of watching someone eyeball stats on their model as they were tweaking parameters, trying to tamp down overfitting.
it's also just such shitty science. "can we find some way to represent this data to conform to $x hypothesis?", albeit that of course isn't surprising from the P-Hacking As A Service crowd
blakestacey@awful.systems · 14 pts · 2y
As an AI language model, I'd like to point everyone to Max Tegmark's appearances in the old!sneerclub archives.
self@awful.systems · 12 pts · 2y
as a large language model, I am incapable of feeling surprise that Tegmark is associated with neo-nazis
(also I really need to de-jank the stylesheet for the archive and get the rest of the data in it soon)
froztbyte@awful.systems · 12 pts · 2y
some of these replies (those are diff links) are staggeringly awful
and this one is a piece of art:
zogwarg@awful.systems · 11 pts · 2y
^^ Quietly progressing from humans are not the only ones able to do true learning, to machines are the only ones capable of true learning.
Poetic.
PS: Eek at the *cough* extrapolation rules lawyering 😬.
swlabr@awful.systems · 10 pts · 2y
Oof, they got so close on that last one, yet so far away. Truly a masterpiece in misunderstanding
carlitoscohones@awful.systems · 11 pts · 2y
My first thought in watching the animation was - AI parrot shoots shotgun at side of barn, draws target around result.
Are the dots in the ocean from Amazon's underwater warehouse structures?
swlabr@awful.systems · 10 pts · 2y
The ocean dots are Lemurian and Atlantean civilisations, situated in the deep. The AI said they’re there, so they must be real! That’s how reality works now!
pikesley@mastodon.me.uk · 5 pts · 2y
@swlabr @carlitoscohones just off to raise some VC millions for my AI-Driven Atlantis Recovery startup brb
200fifty@awful.systems · 10 pts · 2y
I had the same thought as Emily Bender's first one there, lol. The map is interesting to me, but mostly as a demonstration of how anglosphere-centric these models are!
Kichae@kbin.social · 8 pts · 2y
So, what's going on here, in plainer language, anyway? Are they just including location information in training data and then, totally surprisingly finding it again in the output data? That's kind of the sense I get from the post here, but I'm not sure if I'm misunderstanding.
Or did they just cluster the data and squint until someone said one of the graphs "kinda looks like it lines up with a Mercator projection"?
froztbyte@awful.systems · 16 pts · 2y
I.... so. damn you, I looked.
this says
their code does... a lot of things with that input data. including filling some in and conveniently removing "small" towns and some states and eliminating duplicates[1] and other shit
a very quick glance at some of the input data:
cool. so. we have high-precision data with actual coordinates and well-defined information. as the input. to the mash-things-together-into-a-proximates-slurry machine.
and then on prompting the slurry with questions about "hey where is Wyoming", it can provide a rough answer.
amazing.
[1] - whoops forgot the footnote. how about that Washingon in every state, huh? sure is a good thing the US doesn't have lots of reused names!
Kichae@kbin.social · 10 pts · 2y
Wow. I was kinda tongue-in-cheeking it there, because I genuinely thought I was misinterpreting/over-simplifying the OP, but they really are trying to sell "it didn't discard this data we explicitly fed it" as some kind of big deal.
I was expecting this to be more like them discovering that regional dialects exist or soemthing dumb-but-not-that-dumb.
froztbyte@awful.systems · 7 pts · 2y
promptfans, making grandiose badfaith claims that turn out so not-even-wrong it entirely moves the goalposts on the argument? nevarrrr
froztbyte@awful.systems · 8 pts · 2y
these fucking people
self@awful.systems · 7 pts · 2y
we need a run of @dgerard@awful.systems’s “it can’t be that stupid, you must be explaining it wrong” stickers but with the ChatGPT logo instead of the bitcoin one
also how can we talk shit about LLMs when computation was impossible until they were invented?
blakestacey@awful.systems · 10 pts · 2y
15-ish years ago, I was doing a lot of principal component analysis and multi-dimensional scaling. A standard exercise in that area is to take distances between cities, like the lengths of airline flight paths, and reconstruct a map. If only I'd thought to claim that to be a world model!
blakestacey@awful.systems · 7 pts · 2y
Fucking Christ, that hurt to read.
swlabr@awful.systems · 5 pts · 2y
Yes
gerikson@awful.systems · 8 pts · 2y
I'm sorry, I can't take anyone with a blue checkmark seriously.
Kichae@kbin.social · 1 pts · 2y
swlabr@awful.systems · 1 pts · 2y
Also yes
zogwarg@awful.systems · 0 pts · 2y