MaiNLP lab, LMU Munich

@mainlp.bsky.social

MaiNLP research lab at CIS, LMU Munich directed by Barbara Plank @barbaraplank.bsky.social Natural Language Processing | Artificial Intelligence | Computational Linguistics | Human-centric NLP

Super excited to present our work at #ACL2026 ✨ 👉 LLM-based social simulations have lots of analytic flexibility—response generation methods are often overlooked 📍Tue July 7, 9am, Poster Session F

Georg Ahnert @IC2S2@ahnert.eurosky.social · 3mo ago

Thrilled to share that our paper with @carohaensch.bsky.social, @barbaraplank.bsky.social & @mstrohm.bsky.social has been accepted to #ACL2026 Main! How to generate closed-ended survey responses 📊 with LLMs trained to produce open-ended text? 📝 We show results from 32 mio. simulated responses… 🧵

A scientific paper titled "Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models", written by Georg Ahnert, Anna-Carolina Haensch, Barbara Plank, and Markus Strohmaier. Additionally, a central figure of the paper with specification curves is shown.

Excited for 5 papers at #ACL2026NLP with my group and with collaborators. 📍 You can find the work here: 🗓️ Sun. July 5 AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages Oral Session B: Multilinguality and Language Diversity 2

Bild

𝗚𝗿𝗮𝗱𝗶𝗲𝗻𝘁 𝗱𝗲𝘀𝗰𝗲𝗻𝘁 𝗶𝘀 𝗮𝗹𝗹 𝘆𝗼𝘂 𝗻𝗲𝗲𝗱 𝗳𝗼𝗿 𝗶𝗻𝘁𝗲𝗿𝗽𝗿𝗲𝘁𝗮𝗯𝗶𝗹𝗶𝘁𝘆? In our ICML paper, we show that it might just be. ExPLAIND is a method that unifies data attribution, model component attribution, and training dynamics by computing an exact decomposition of model behavior based on gradient products.

BildBild

In our new paper, "A Comprehensive Evaluation of Multilingual Chain-of-Thought Reasoning: Performance, Consistency, and Faithfulness Across Languages", we go beyond final-answer accuracy to analyze multilingual reasoning along three dimensions: performance, consistency, and faithfulness.

✨New paper✨ We find script (e.g. Cyrillic, Latin) to be a linear direction in the activation space of Whisper, enabling transliteration at test-time by adding such script directions to the activations — producing e.g. Cyrillic Japanese transcriptions.

Bild

At #Interspeech2025 I'm going to present Betthupferl, a dataset for German dialect ASR & dialect-to-standard speech translation! We analyze differences between dialectal & Standard German transcriptions, benchmark ASR models, and examine shortcomings of current ASR models & evaluation metrics.

Piper title ("A multi-dialectal dataset for German dialect ASR and dialect-to-standard speech translation") and a map of the German state Bavaria showing where the Franconian, Bavarian, and Alemannic dialect groups are spoken

UPDATE: Our poster presentation got moved to Tuesday, 16:00–17:30 (session 10)! #ACL2025NLP

Verena Blaschke@verenablaschke.bsky.social · last yr.

At #ACL2025NLP I'll present our analysis of the effect of linguistic similarity on cross-lingual transfer! We looked at how 10 similarity measures correlate w/ transfer results btwn 263 languages across 3 NLP tasks. Different similarity measures matter for diff. experiments (no one-size-fits-all)!

Correlations between transfer results per experiment (parsing, POS tagging, topic classification with different input representations) and similarity measures. The results vary a lot across experiments and measures – some are described in the next posts.

At #ACL2025NLP I'll present our analysis of the effect of linguistic similarity on cross-lingual transfer! We looked at how 10 similarity measures correlate w/ transfer results btwn 263 languages across 3 NLP tasks. Different similarity measures matter for diff. experiments (no one-size-fits-all)!

Correlations between transfer results per experiment (parsing, POS tagging, topic classification with different input representations) and similarity measures. The results vary a lot across experiments and measures – some are described in the next posts.