Mapping the Mind of a Large Language Model

kromem@lemmy.world · 6 months ago

Mapping the Mind of a Large Language Model

Womble@lemmy.world · edit-2 6 months ago

This is a really good science communication article, it describes their work in clear terms (finding structures that relate to abstract concepts, seeing when they are activated and how strengthening and weaking them modifies outputs) and goes into the implications for it. I’m probably going to save this link as a rebuttal for the people who claim LLMs just predict the next word and have no concepts embedded in them.

misk@sopuli.xyz · 6 months ago

I doubt that anyone saying that LLM are calculating next word solely based on previous sequence. It’s still statistics, regardless of complexity.

Richard@lemmy.world · 6 months ago

Yes, but people forget that our brains, and therefore our minds, are also “simply” statistics, albeit very complex.

Womble@lemmy.world · 6 months ago

Youd be surprised at the level of unthinking hatred around them, but even discarding that Ive seen it said often that LLMs have no internal model of what they are talking about as they are just next word generators. This quite clearly contradicts that interpretation.

technocrit@lemmy.dbzer0.com · edit-2 6 months ago

There is no mind. It’s pretty clear that these people don’t understand their own models. Pretending that there’s a mind and the other absurd anthropomorphisms doesn’t inspire any confidence. Claude is not a person jfc.

Drewelite@lemmynsfw.com · edit-2 6 months ago

Ah yes, it must be the scientists specializing in machine learning studying the model full time who don’t understand it.