Maxime Méloux

@maximemeloux.bsky.social

PhD student @LIG | Causal abstraction, interpretability & LLMs

Our new paper, "The Dead Salmons of AI interpretability", is out! In 2009, researchers showed that standard statistical errors could detect "brain activity" in a dead salmon 🐟. Modern XAI methods face similar issues: we find interpretable neurons and probes even in randomly initialized models. 1/X