Institute of Formal and Applied Linguistics

@ufal.mff.cuni.cz

Computational linguistics • Natural language processing • Formal linguistics • Machine translation | at Faculty of Mathematics and Physics, Charles University

Congrats! Stop by to chat with @ivankartac.bsky.social at #ACL2026 in San Diego 🇺🇸 He's presenting 2 papers: A modular neuro-symbolic system pairing small LLMs with a theorem prover for syllogistic reasoning, and BOULDER, a new benchmark showing LLM reasoning drops sharply inside multi-turn dialogue.

Ivan Kartáč@ivankartac.bsky.social · last mo.

Heading to San Diego for #ACL2026, where I’ll be presenting two papers (see 🧵). Stop by to chat about evaluating reasoning embedded in task-oriented dialogue, or how to use small LLMs in modular neuro-symbolic approaches to syllogistic reasoning!

If you are at #ACL2026, don't miss @mariemikulova.bsky.social et al. presenting a semantic-pragmatic annotation layer for the 3M+ token Prague Dependency Treebank aclanthology.org/2026.finding...

aclanthology.org

Marie Mikulová@mariemikulova.bsky.social · last mo.

Semantic-pragmatic annotation in the Prague Dependency Treebank — three times at #ACL2026! 📍 Jul 3, 13:30–14:15 LAW La Jolla, talk 📍 Jul 3, 14:20–15:30 CODI-CRAC Regatta, poster 📍 Jul 6, 11–12:30 Findings Poster, Grand Hall See you in San Diego! @ufal.mff.cuni.cz aclanthology.org/2026.finding...

🔸 🔶 Petr Sgall was a pioneer who saw the future of NLP long before the age of LLMs 🔶 🔸 As we mark the 100th anniversary of his birth, Eva Hajičová and Jarmila Panevová have shared a brilliant tribute to his life and work. 🔗 Read the full tribute here: www.matfyz.cz/clanky/ste-v...

Sté výročí průkopníka počítačové lingvistiky prof. Petra Sgalla

V květnu tohoto roku si připomínáme 100 let od narození emeritního profesora Univerzity Karlovy a dlouholetého pracovníka Matematicko-fyzikální fakulty UK PhDr. Petra Sgalla, Dr. h. c. mult. (27. květ...

matfyz.cz

New work from our team on morphology-aware tokenization evaluation, accepted to EACL 2026 Findings.

Jindřich Libovický@jlibovicky.bsky.social · 6mo ago

We (= mostly @abyste.bsky.social) developed a way to evaluate how morphological a #tokenization is w/o gold segmentation labels. arxiv.org/abs/2601.18536 The key: align subword tokens with morphological features from UniMorph using IBM Model 1. To appear in EACL 2026 Findings.

If you think labeling text spans with LLMs is easy, you probably have not tried it yourself (we have! 🙃). Any method you can think of – be it tagging, matching, or indexing – has flaws. In our new preprint, we tested them all 💪We also proposed how to improve one of them. arxiv.org/abs/2601.16946

Bild