Mastodawn

kromem May 21, 2024

Mapping the Mind of a Large Language Model

Mapping the Mind of a Large Language Model - Lemmy.World

I often see a lot of people with outdated understanding of modern LLMs. This is probably the best interpretability research to date, by the leading interpretability research team. It’s worth a read if you want a peek behind the curtain on modern models.

Show thread

Womble May 22, 2024

This is a really good science communication article, it describes their work in clear terms (finding structures that relate to abstract concepts, seeing when they are activated and how strengthening and weaking them modifies outputs) and goes intobthe implications for it. Im probably going to save this link as a rebuttal for the people who claim LLMs just predicr the next word and have no concepts embedded in them.

Show thread

misk May 22, 2024

I doubt that anyone saying that LLM are calculating next word solely based on previous sequence. It’s still statistics, regardless of complexity.

Show thread

ricdeh May 22, 2024

Yes, but people forget that our brains, and therefore our minds, are also “simply” statistics, albeit very complex.

Show thread

Zos_Kia

Yeah I found this kind of reductionist talk pushes people to overlook the emerging properties of the system, which is where the meat of the topic is. It’s like looking at a living cell and saying “yeah well this is just chemistry”.