Sign in

Maxime Méloux

@maximemeloux.bsky.social
78 followers 191 following 9 posts

PhD student @LIG | Causal abstraction, interpretability & LLMs

PostsRepliesMedia
Maxime Méloux @maximemeloux.bsky.social · 10/07/2026
I'm very happy to give a spotlight at the Mechanistic Interpretability Workshop @ ICML on our work: "Validating Causal Abstraction Metrics on Simulated Complex Systems" Which metrics actually tell you if an explanation is valid? We built a benchmark to find out. 1/n
melouxm.github.io
Validating Causal Abstraction Metrics on Simulated Complex Systems
142
Maxime Méloux @maximemeloux.bsky.social · 08/01/2026
Our new paper, "The Dead Salmons of AI interpretability", is out! In 2009, researchers showed that standard statistical errors could detect "brain activity" in a dead salmon 🐟. Modern XAI methods face similar issues: we find interpretable neurons and probes even in randomly initialized models. 1/X
130
Maxime Méloux @maximemeloux.bsky.social · 26/04/2025
I'm very happy to present our work "Everything, Everywhere, All at Once: Is Mechanistic Interpretability Identifiable?" this afternoon at #ICLR2025! Come have a chat at stand #439 :)
0101