"Mechanistic indicators of Understanding in LLMs" is finally out in Philosophical studies!
link.springer.com/article/10.1...
link.springer.com
Mechanistic indicators of understanding in large language models - Philosophical Studies
Philosophical Studies - Large language models are often portrayed as merely imitating linguistic patterns without genuine understanding. We argue that recent findings in mechanistic...