Verena Blaschke

@verenablaschke.bsky.social

Postdoc @gronlp.bsky.social (University of Groningen), interested in language variation, NLP for low-resource varieties, and in fair & ethical language technology development. verenablaschke.github.io

At this year's ACL business meeting, it was wonderful to see so many attendees - a sign that this community cares. As president, I wanted to use the occasion to give some insights into ACL - the premier organization for computational linguistics, founded in 1962. (1/5)

BildBild

Interested in developing LLMs that work for dialectal Arabic? Introducing the AMIYA shared task: Arabic Modeling In Your Accent, just accepted to VarDial 2026. Please consider submitting and joining us in Morocco if you do! sites.google.com/view/vardial...

VarDial 2026 - Shared Tasks

AMIYA (عامية) Shared Task: Arabic Modeling In Your Accent The AMIYA shared task will offer a chance for researchers to demonstrate innovations and improvements in language modeling of dialectal Arabic...

sites.google.com

At #Interspeech2025 I'm going to present Betthupferl, a dataset for German dialect ASR & dialect-to-standard speech translation! We analyze differences between dialectal & Standard German transcriptions, benchmark ASR models, and examine shortcomings of current ASR models & evaluation metrics.

Piper title ("A multi-dialectal dataset for German dialect ASR and dialect-to-standard speech translation") and a map of the German state Bavaria showing where the Franconian, Bavarian, and Alemannic dialect groups are spoken

UPDATE: Our poster presentation got moved to Tuesday, 16:00–17:30 (session 10)! #ACL2025NLP

Verena Blaschke@verenablaschke.bsky.social · last yr.

At #ACL2025NLP I'll present our analysis of the effect of linguistic similarity on cross-lingual transfer! We looked at how 10 similarity measures correlate w/ transfer results btwn 263 languages across 3 NLP tasks. Different similarity measures matter for diff. experiments (no one-size-fits-all)!

Correlations between transfer results per experiment (parsing, POS tagging, topic classification with different input representations) and similarity measures. The results vary a lot across experiments and measures – some are described in the next posts.

At #ACL2025NLP I'll present our analysis of the effect of linguistic similarity on cross-lingual transfer! We looked at how 10 similarity measures correlate w/ transfer results btwn 263 languages across 3 NLP tasks. Different similarity measures matter for diff. experiments (no one-size-fits-all)!

Correlations between transfer results per experiment (parsing, POS tagging, topic classification with different input representations) and similarity measures. The results vary a lot across experiments and measures – some are described in the next posts.

Dei Boarisch heard ned bei "Servus" und "Pfiade" auf? Dann suach ma genau Di! Wir suachan Bairischsprecher:innen, de a kurze Umfrage über KI-generierds Boarisch für a Masterarbeit beantwortn mechadn. Mid jeder Teilnahme bring ma den boarischn Dialekt a Stickal weida in de digitale Weyd!

Verena Blaschke@verenablaschke.bsky.social · last yr.

Bavarian dialect speakers needed! Our MSc student Miriam wants to find out 1. how good/bad LLM-generated "Bavarian" is, and 2. whether dialect speakers agree with each other on this. The survey takes <5 min: survey.ifkw.lmu.de/dialquali25/ Thank you for sharing/participating!

The first archival *CL Queer in AI workshop will kick off in about 15 min! Join us in-person if you're at NAACL or virtually 💜 We will have presentations from our amazing contributors and invited speakers. Read on for more details 🧵

On my way to #NAACL2025 where I'll give a keynote at the noisy text workshop (WNUT), presenting some of the challenges & methods for dialect NLP + also discussing dialect speakers' perspectives! 🗨️ Beyond “noisy” text: How (and why) to process dialect data 🗓️ Saturday, May 3, 9:30–10:30