At EACL '26, TakeLab members are presenting two papers. Feel free to stop by if you'd like to chat about simple tricks to improve LLM token embeddings, principled evaluation of synthetic data, coffee, or life in general. Details below 👇🧵 @eaclmeeting.bsky.social #EACL2026 #NLProc 🇲🇦
David Dukić
@ddaviddukic.bsky.social
PhD in NLP | TakeLab 🇭🇷 | Information extraction, representation learning & analysis | Making LLMs better one step at a time
We already know prompt repetition is a handy hack to improve a decoder-only LM’s performance as it allows the model to “see” bidirectionally, an ability otherwise suppressed by the causal mask. But what happens if we increase the number of repetitions? 🤔🧵 @eaclmeeting.bsky.social #EACL2026
📣📣 New preprint alert!! Despite events in the world becoming bleaker, the news is… more positive? We conduct a diachronic study of word embeddings trained on 10M Croatian news articles spanning 25 years and find some surprising results! arxiv.org/abs/2506.13569