@a-krishnan.bsky.social

Master student at Saarland university

Back from Seoul. My first paper, "An Isotropic Approach to Efficient UQ with Gradient Norms", got a poster and an oral at ProbML, and came away with the Best Paper Award. Still a bit stunned. arxiv.org/abs/2603.29466

Bild

📢 #SpeechTech & #SpeechScience researchers! We are thrilled to announce that Prof. Karen Livescu will keynote our Special Session on Interpretable Audio and Speech Models at #Interspeech2025: "What can interpretability do for us (and what can it not)?" 🗓️ Aug 18, 11:00 @interspeech.bsky.social

Announcements

Keynote Speaker Announcement 🔊 30.07.2025 We are delighted to announce the keynote speech t`hat will happen at the special session! Speaker: Prof. Karen Livescu, Toyota Technological Institute at Ch...

sites.google.com

Chain-of-Thought (CoT) reasoning lets LLMs solve complex tasks, but long CoTs are expensive. How short can they be while still working? Our new ICML paper tackles this foundational question.

Bild

AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories We are releasing the first benchmark to evaluate how well automatic evaluators, such as LLM judges, can evaluate web agent trajectories.

Bild

Checkout Benno's notes about our impact of interpretability paper 👇. Also, we are organizing a workshop at #ICML2025 which is inspired by some of the questions discussed in the paper: actionable-interpretability.github.io

General Information

ICML 2025 - Vancouver

actionable-interpretability.github.io

Benno Krojer@bennokrojer.bsky.social · last yr.

Day 12: From Insights to Actions: The Impact of Interpretability and Analysis Research on NLP arxiv.org/abs/2406.12618 Genuinely one of my favourite papers in recent years! It tries to answer one question that every phd student often asks themselves: Does this research matter?

Agents like OpenAI Operator can solve complex computer tasks, but what happens when users use them to cause harm, e.g. spread misinformation? To find out, we introduce SafeArena (safearena.github.io), a benchmark to assess the capabilities of web agents to complete harmful web tasks. A thread 👇

Bild