With a high volume of submissions this year, we're recruiting additional reviewers for BlackboxNLP 2026! 🗓 Review deadline: August 17 (AoE) 🗓 Review load: 2-4 papers 📝 Sign up: forms.gle/tha166UQZYqX...
Martin Tutek
@mtutek.bsky.social
Postdoc @ TakeLab, UniZG | previously: Technion; TU Darmstadt | PhD @ TakeLab, UniZG Faithful explainability, controllability & safety of LLMs. 🔎 On the academic job market 🔎 https://mttk.github.io/
For today's reading group, @veraneplenbroek.bsky.social presented "Old Habits Die Hard: How Conversational History Geometrically Traps LLMs" by Simhi et al. (2026) Paper: arxiv.org/abs/2603.03308 #NLProc
⏳ The BlackboxNLP 2026 Reproducibility Challenge deadline has been extended to July 24 (AoE) ⏳ If you've been working on a robustness check, ablation, or replication of recent NLP interpretability work, you have a bit more time to get your submission in.
⏳ One week left! The submission deadline for BlackboxNLP 2026 is July 17 (AoE). If you're working on analyzing or interpreting neural networks for NLP, now's the time to get your paper in. 📍 Co-located with #EMNLP2026 in Budapest 🔗 blackboxnlp.github.io
🏆 Announcing the NDIF Best Paper Award for the BlackboxNLP 2026 Reproducibility Challenge! @ndif-team.bsky.social Reproduce an interp. finding with nnsight + NDIF, open-source it, and push it further. Winners will receive a $500 prize, and will be invited to present at the workshop!
At ICML '26, TakeLab members are presenting three papers. Make sure to stop by if interested in language model interpretability, coding agent evaluations, or pluralistic alignment. Details below 👇🧵 @icmlconf.bsky.social #ICML2026 🇰🇷
📣 #EACL2027 updates: the Call for Papers is live, and our keynote speakers are confirmed! Main conference: 9–14 March 2027. Special theme: The Human in Language. 🧵👇 2027.eacl.org/calls/papers/
Call for Papers
Official website for the 2027 Conference of the European Chapter of the Association for Computational Linguistics
2027.eacl.org
We don’t always know what problems are hard for LLMs. So devs evaluate on tasks HUMANS find hard or on broad benchmarks. What if we could instead anticipate which scenarios a model will fail on—all without evaluating specific input examples? 🧵NEW PAPER by @jenniferlumeng.bsky.social
✨ it's coming ✨ NEMI 2026 will be lit. It will also be the new BU interp supergroup's debut ball. Come meet us!
The 3rd New England Mechanistic Interpretability (NEMI) Workshop
nemiconf.github.io
With the large influx of submissions and a faster pace of research, reproducibility is more important than ever. With this reproducibility challenge, we want to put the focus on best practices wrt. baselines🧱, ablations🌈, eval🔎 and generalizability🗺️ of interpretability!
📣 Announcing the BlackboxNLP 2026 Reproducibility Challenge! A new track dedicated to rigorous robustness checks of NLP interpretability work - stress-testing baselines, ablations, generalizability, and evaluation.
📣 Announcing the BlackboxNLP 2026 Reproducibility Challenge! A new track dedicated to rigorous robustness checks of NLP interpretability work - stress-testing baselines, ablations, generalizability, and evaluation.
Ever used a top-ranked LLM that just... felt wrong for you? You’re not alone. Instead of leaderboards, many of us turn to "vibe-testing" - manually comparing models to our own needs. But can we turn these feelings into a structured evaluation? New paper: "From Feelings to Metrics" 🧵
We are delighted to welcome @marlutz.bsky.social to our lab over the next few months! 🎉 She'll work on the representation of different demographic groups in LLMs. #NLProc
FYI #ACL2026 has an unusual registration system this year, and probably a lot of people who want to attend will not be able to. Spots are limited to 3.5k people, and only presenting authors can register during the first phase. Then, *if* there are spots left, others can try to register.
Registration
Official website for the 64th Annual Meeting of the Association for Computational Linguistics
2026.aclweb.org
📢 The workshop on Insights from negative results will be back at EMNLP'26! Your most-insightful failures can be submitted in 4 pages by June 25. It's also possible to commit short papers reviewed through ARR. insights-workshop.github.io/2026/cfp
2026 Call for Papers
Workshop on Insights from Negative Results in NLP
insights-workshop.github.io
How can generative AI better support human creativity, without limiting it? If you have thoughts, we invite submissions to our ICML workshop on Generative AI, Creativity, and Human-AI Co-Creation 📍 July 2026, Seoul 📄 Submit by: April 24 (AOE) 🔗 Submission link: openreview.net/group?id=ICM...
ICML 2026 Workshop GenAICreativity
Welcome to the OpenReview homepage for ICML 2026 Workshop GenAICreativity
openreview.net
❗The full paper submission deadline for COLM is ~14 hours from now (11:59pm AOE)! Please submit your final PDFs on the same page where you uploaded your abstracts. And please use the provided LaTeX templates; do not handwrite your manuscript like this llama is! Good luck!
Interested in pursuing a PhD in NLP/cog-sci? Studying language learning in LMs from the perspective of human language acquisition? Few more days to apply!!
📢 PhD position in Developmental Language Modelling (PLZ RT) What can human language acquisition teach us about training language models? Join us as a PhD! mpi.nl/career-education/vacancies/vacancy/fully-funded-4-year-phd-position-developmental-language @carorowland.bsky.social @mpi-nl.bsky.social
A piece co-authored by an old friend (Divya Saini, a psychiatrist at Massachusetts General Hospital) www.nytimes.com/2026/03/29/o...
Opinion | Your Chatbot Isn’t a Therapist
nytimes.com
Check out works on sequence repetition 🔁 and evaluating synthetic data 🧮 from our lab in Rabat! @eaclmeeting.bsky.social #EACL2026
At EACL '26, TakeLab members are presenting two papers. Feel free to stop by if you'd like to chat about simple tricks to improve LLM token embeddings, principled evaluation of synthetic data, coffee, or life in general. Details below 👇🧵 @eaclmeeting.bsky.social #EACL2026 #NLProc 🇲🇦
To ease my FOMO from not attending @eaclmeeting.bsky.social, I skimmed the proceedings while playing the Tangier episode of Parts Unknown. I'll do something different and shout out 5 (subjectively) interesting works of authors I'm *not* closely related to, in no specific order:🧵⬇️
Thinking of applying for an #MSCA Postdoctoral Fellowship in 2026? I’m open to supervising at Bocconi! Feel free to reach out. By submitting an expression of interest to Bocconi, selected applicants will receive full proposal support. 🗓️ Deadline: April 15 👉 www.unibocconi.it/en/horizon-e...
Horizon Europe - Marie Skłodowska-Curie Actions Postdoctoral Fellowships 2026 - Bocconi University
unibocconi.it
Excited to present this work together with @dippedrusk.com at #EACL. Join us in the poster session 1 (11:30-13:00) 🔥
LMs that "know more" about toxicity are less toxic! Our #TACL 📄 connects behavior and internals: 💠 LMs amplify toxicity beyond humans 💠 Information about toxicity peaks in lower layers 💠 Bypassing these layers increases toxicity More details👇 #NLProc #interpretability (1/🧵)
Excited to share that @milanlp.bsky.social will be presenting 5 new papers at #EACL2026 and workshops in Rabat 🇲🇦!
Argh this sucks. apparently they lost their funding guarantee
You are #EACL2026? Check out some great work from the CSS Department @gesis.org, our data Science Methods team, with great collaborators!
I’m seeing close to zero reaction/conversation about this on here. This is huge news for open research on language models, especially in the US.
Wow a big hit for AI2 and for public interest, open source AI research
“It’s toddler AI misinformation at an industrial scale. It’s very risky for the developing brain.” Children’s media experts say AI-generated “slop" has infiltrated the internet, preying on young children and their unsuspecting caregivers.
This is your kid's brain on AI slop
“It’s toddler AI misinformation at an industrial scale. It’s very risky for the developing brain.”
motherjones.com
We'll have a reproducibility track at this years' Blackbox workshop! Details are still within a slightly opaque box. We want to see if cleaning solutions that make opaque boxes 📦 transparent 🍱 work on different boxes 🎁📮🧰🥡, and with different 🧽solution-to-water🧼ratios!
BlackboxNLP will be co-located with EMNLP 2026 in 🇭🇺 Budapest 🇭🇺 this October! This edition will feature a special reproducibility track, investigating generalization and robustness of established results from interpretability research 👷♂️ Stay tuned for more details!