Sign in

Millicent Li

@millicentli.bsky.social
21 followers 13 following 8 posts

CS PhD Student @ Northeastern, former ugrad @ UW, UWNLP -- millicentli.github.io

PostsRepliesMedia
Reposted by Millicent Li
Aaron Mueller @amuuueller.bsky.social · 01/10/2025
What's the right unit of analysis for understanding LLM internals? We explore in our mech interp survey (a major update from our 2024 ms). We’ve added more recent work and more immediately actionable directions for future work. Now published in Computational Linguistics!
24115
Millicent Li @millicentli.bsky.social · 17/09/2025
Wouldn’t it be great to have questions about LM internals answered in plain English? That’s the promise of verbalization interpretability. Unfortunately, our new paper shows that evaluating these methods is nuanced—and verbalizers might not tell us what we hope they do. 🧵👇1/8
1268