Antonin Poché

@antoninpoche.bsky.social

PhD Student doing XAI for NLP at @ANITI_Toulouse, IRIT, and IRT Saint Exupery. 🛠️ Interpreto & Xplique library development team member. https://antoninpoche.github.io/

Tomorrow (Monday, 6th), I am presenting Interpreto at ACL San Diego! 🛬 Interpreto is an open-source library for interpreting language models (attributions, probes, SAEs...). I will be in Demo Session E: Grand Hall from 11 am to 1 pm. 📺Look for the TVs, and you will find me!

Bild

Yesterday, with @fannyjrd.bsky.social we were delighted to share how explainability can be used for fairness in a @facct.bsky.social tutorial!! We loved the questions and interactions! Check Fanny's thread for the link. The conference is really fun thus far! Thanks to the organizers!

Fanny Jourdan@fannyjrd.bsky.social · last mo.

Today, with @antoninpoche.bsky.social, we had the pleasure of presenting a tutorial at @facct.bsky.social on how explainability can be used as a practical tool for fairness analysis. 1/4

BlackboxNLP will be co-located with EMNLP 2026 in 🇭🇺 Budapest 🇭🇺 this October! This edition will feature a special reproducibility track, investigating generalization and robustness of established results from interpretability research 👷‍♂️ Stay tuned for more details!

Bild

🔥I am super excited for the official release of an open-source library we've been working on for about a year! 🪄interpreto is an interpretability toolbox for HF language models🤗. In both generation and classification! Why do you need it, and for what? 1/8 (links at the end)

Bild

🕳️🐇 𝙄𝙣𝙩𝙤 𝙩𝙝𝙚 𝙍𝙖𝙗𝙗𝙞𝙩 𝙃𝙪𝙡𝙡 – 𝙋𝙖𝙧𝙩 𝙄 (𝑃𝑎𝑟𝑡 𝐼𝐼 𝑡𝑜𝑚𝑜𝑟𝑟𝑜𝑤) 𝗔𝗻 𝗶𝗻𝘁𝗲𝗿𝗽𝗿𝗲𝘁𝗮𝗯𝗶𝗹𝗶𝘁𝘆 𝗱𝗲𝗲𝗽 𝗱𝗶𝘃𝗲 𝗶𝗻𝘁𝗼 𝗗𝗜𝗡𝗢𝘃𝟮, one of vision’s most important foundation models. And today is Part I, buckle up, we're exploring some of its most charming features. :)

🔥 I am super excited to be presenting a poster at #ACL2025 in Vienna next week! 🌏 This is my first big conference! 📅 Tuesday morning, 10:30–12:00, during Poster Session 2. 💬 If you're around, feel free to message me. I would be happy to connect, chat, or have a drink!

Bild

🚨 New preprint! 🚨 Everyone loves causal interp. It’s coherently defined! It makes testable predictions about mechanistic interventions! But what if we had a different objective: predicting model behavior not under mechanistic interventions, but on unseen input data?

Bild

🔥ConSim has been accepted to the #ACL2025 main conference! 🙏 Thanks again to my amazing co-authors: @alon_jacovi, Agustin Picard, @VictorBoutin, and @Fannyjrd_. Work done in DEEL and FOR from IRT St Exupéry and @ANITI_Toulouse. See you in Vienna 📅 For more information, check out my last post:

Antonin Poché@antoninpoche.bsky.social · 2y ago

🚀 Thrilled to share our new paper (the first of my PhD)! How can we compare concept-based #XAI methods in #NLProc? ConSim (arxiv.org/abs/2501.05855) provides the answer. Read the thread to find out which method is the most interpretable! 🧵1/7

An assembly of 18 European companies, labs, and universities have banded together to launch 🇪🇺 EuroBERT! It's a state-of-the-art multilingual encoder for 15 European languages, designed to be finetuned for retrieval, classification, etc. Details in 🧵

Bild