Ivan Kartáč

@ivankartac.bsky.social

PhD student @ Charles University, Prague. NLP & computational linguistics. Working on evaluation, explainability, and reasoning. ivankartac.github.io

It has been *wild* to me to see the way that people in my department have fully gone back to business-as-usual posting on Twitter for their papers. There were a couple of brief blips where people tried bluesky and LinkedIn, but that's nearly all gone now 😔

I recommend this article about AI reasoning, where the author lets us in on his struggles w/ AI cognitive dissonance. Plus some priceless quotes from @rao2z.bsky.social. (My recommendation has *nothing* to do with the fact that I'm quoted in it too 😇) www.quantamagazine.org/is-ai-reason...

Is AI Reasoning Right for the Wrong Reasons? | Quanta Magazine

The idea that artificial intelligence can “reason” is more intuitive than ever. But intuitions can be wrong, and the science is far from settled.

quantamagazine.org

We need a benchmark where systems generate recipes and people cook based on it and rate the results. Item Response Theory will handle annotators with poor cooking skills and too easy or infeasible meals.

I resigned from Google DeepMind bc it broke its founding promise by selling AI to the military without restrictions against killer robots or mass spying. For months, I worked to stop this but watched powerful ethicists and institutions choose silence. Here's what happened. 🧵

Bild

My assumption has been that in academia, humanities are much less inclined to use GenAI for their work. To what extent is this true? Or are people in humanities just less willing to acknowledge the use?

I think I will be posting this after each #ARR cycle: Please 🥺🙏 let's prohibit AI review writing. Otherwise, we will get lazy reviewers claiming they wrote bullet points and used AI only to put that in prose.

Heading to San Diego for #ACL2026, where I’ll be presenting two papers (see 🧵). Stop by to chat about evaluating reasoning embedded in task-oriented dialogue, or how to use small LLMs in modular neuro-symbolic approaches to syllogistic reasoning!

I'd never have guessed models commit to their final answer this early, often within the first 20% of reasoning, across math/logic tasks and model families. The rest is mostly hedging that doesn't change their mind. And turns out they encode this internally, we can decode it! 🧵👇

Sara Candussio@saracandussio.bsky.social · last mo.

Are all the CoT steps necessary? In our latest paper, we find evidence for the existence of a commitment boundary, marking a sharp transition from no/mid guesses to the model final answer across various reasoning tasks and model families. Thread 🧵👇

“Dimicillin” isn’t real. We made it up. Yet many LLMs still call it an antibiotic. Across 9 models and 653 drugs, we find that drug-name affixes alone can drive pharmacological reasoning. Models often rely on morphology over facts. We trace this shortcut from behavior to mechanism. 🧵

BildBild

Do you sometimes have to explain to engineers that the main role of science is not to produce software? Once in a while I see people comment on some paper along the lines of “but it’s not efficient” or “I can’t use this in production” as if this was what research is about.

New blog: I am worried by NLP research culture NLG and NLP are mostly much better in 2026 than when I got my PhD in 1990. Unfortunately research culture has gotten *worse” in this period, which really worries me as I retire. ehudreiter.com/2026/06/08/n...

I am worried by NLP research culture

In most ways NLG and NLP are much better in 2026 than when I got my PhD in 1990. Unfortunately research culture has gotten *worse” in this period, which really worries me as I retire. We have…

ehudreiter.com

Maybe one of the biggest obstacles for progress in science comes from entrenched stereotypes? In linguistics, we have, for example, (1) the word stereotype, (2) the grammar/dictionary stereotype, (3) the building-block stereotype, and (4) the speaker directionality stereotype dlc.hypotheses.org/4343

Four stereotypes that have guided morphosyntactic thinking

Thinking about language structures is made difficult not only by their incredible complexity, but also by entrenched ways of thinking about grammatical and lexical patterns. Linguists do not investiga...

dlc.hypotheses.org