@maikezufle.bsky.social
1️⃣ "Do What I Say: A Spoken Prompt Dataset for Instruction-Following" 👥 @maikezufle.bsky.social, @sarapapi.bsky.social, Fabian Retkowski, Szymon Mazurek, Marek Kasztelnik, Alexander Waibel, @luisabentivogli.bsky.social, @jan-niehues.bsky.social 🇪🇺 Meetween EU project 📄 arxiv.org/abs/2603.09881
Do What I Say: A Spoken Prompt Dataset for Instruction-Following
Speech Large Language Models (SLLMs) have rapidly expanded, supporting a wide range of tasks. These models are typically evaluated using text prompts, which may not reflect real-world scenarios where ...
arxiv.org
We'll be hosting a tutorial at EAMT (already next week!), KONVENS and MT Marathon on human evaluation. Come learn with us! With support of @maikezufle.bsky.social and @patuchen.bsky.social
Deadline extension! - The task is simple: get audio + its translation and estimate how good it is. - Mark your name as the winner of the first Speech Translation Metrics Shared Task at IWSLT 2026 🏆 Predictions submission: May 7, 2026 Description paper: May 10, 2026
At our last seminar, @maikezufle.bsky.social presented her work on Duplex Models, "Building Controllable Speech Systems"
Excited to see our work presented at EACL 2026! I sadly won’t make it this year to Rabat, but definitely chat with @zouharvi.bsky.social about Early-Exit COMET if you’re attending. 🏃🏻♀️➡️🚪
Quality estimation (automated metrics) are amazing. Truly. We would like to use them everywhere. That gets compute-expensive very quickly. We also don't know when they don't know. In "Early-Exit and Instant Confidence Translation Quality Estimation" (at EACL26) we fix that.
Quality estimation (automated metrics) are amazing. Truly. We would like to use them everywhere. That gets compute-expensive very quickly. We also don't know when they don't know. In "Early-Exit and Instant Confidence Translation Quality Estimation" (at EACL26) we fix that.
Join us at the first ever Speech Translation Metrics Shared Task at IWSLT 2026! ✨✨
Have you ever wondered how speech translation gets evaluated? Sadly, most speech evaluation downgrades to text-based metrics. Let’s do better! At IWSLT 2026, we’re launching the first-ever ✨Speech Translation Metrics Shared Task ✨!
It was great seeing you all at EMNLP in Suzhou. 🤗🌆 I presented two papers, check them out if you are interested in 🎤 Speech (summarization) or 📊 MT evaluation with multiple candidates. aclanthology.org/2025.emnlp-m... aclanthology.org/2025.wmt-1.63/
I am heading to Suzhou 🌆 for EMNLP, presenting two papers with Fabian Retkowski, @zouharvi.bsky.social and Tu Anh Dinh. Looking forward to see you at our posters about 🎤 Speech Summarisation (Fri. Nov 7, 14:00-15:30) and 📊 Machine Translation Evaluation (Sat. Nov 8, 11:00-12:00).
Excited to head over to #Suzhou to present 5 papers at #EMNLP2025 and affiliated venues! Topics include quality estimation and evaluation 👩🔬, speech🗣️, and multilinguality🌐- see you soon! 🤩 @maikezufle.bsky.social @jan-niehues.bsky.social
Ich freue mich schon auf den Vortrag und Austausch!
Wir bei AI4LT arbeiten täglich mit 🤖 KI und 🎤 Sprache. Aber wie funktioniert Künstliche Intelligenz eigentlich – und wie kann sie Texte und gesprochene Sprache erzeugen? @maikezufle.bsky.social aus unserer Gruppe erklärt es in einem Vortrag in der Stadtbücherei #Wörth 📚. English version below👇
🇦🇹First evening at #ACL2025NLP with KIT‘s IWSLT Shared Task team. We ranked 1st overall in 🥇offline (Sai Koneru) 🥇instruction following (Maike Züfle) 🥇Arabic dialects (Zhaolin Li) tracks😍 👋Posters: 🖼️Thur, 14-15:30 (Offl.+ Instr. Fol.) 🖼️Fri, 11-12:30 (Arabic) Papers in 🧵
🇦🇹 I’ll be in Vienna for #ACL2025NLP! Interested in training a speech-aware LLM without a lot of parameteres or data? Come to my poster: 🖼️ Mon, 18:00 Also into Speech Summarization? Join my IWSLT talk in collab with @fbk-mt.bsky.social 🎤 Fri, 14:00 Happy to chat - come say hi! 😎 Papers in 🧵
Excited to share 7 papers accepted at #ACL2025 and affiliated venues! 🎉 Topics include multimodality, cross-lingual transfer, benchmarks, decoding, low-resource languages and more. See you in Vienna! 🤩 #nlproc #acl2025nlp
@zouharvi.bsky.social and I had fun presenting our poster on efficient reranking for machine translation! Check out our paper with @juliuscheng.bsky.social aclanthology.org/2025.naacl-l...
Being in a hot air balloon in Albuquerque really makes one ponder *how to efficiently pick the best translation candidate without running expensive evaluation metrics on all of them.* See you tomorrow at 9:00 in Hall 3 #NAACL2025.
We had a great time hosting enthusiastic high schoolers at #GirlsDay2025! We explored #LLM basics, played "Break the AI" games, and shared study abroad opportunities at #KIT. Thanks @maikezufle.bsky.social for organizing and @kit.edu @girls-day.de for this initiative! 👩💻🌟
Happy to announce that our work "A Bayesian Optimization Approach to Machine Translation" was accepted to NAACL 2025! Special thanks to @ufal-cuni.bsky.social for organizing MT Marathon 2025 where I was able to team up with @maikezufle.bsky.social and @zouharvi.bsky.social ! Explainer below: