Najoung Kim

@najoung.bsky.social

https://najoung.kim langauge

i will be at #ICML2026! first time in my career (!) that I'm attending a major conf in my home town and super excited about it 💖🇰🇷 (but also... clash of worlds 😬) I'll be doing hallway track for the most part so DM or email me if you'd like to hang! + don't forget to 👇

🐸 New position paper on compositionality! 🐸 I synthesize my thoughts about the proper role/interpretation of behavior and mechanism in asking the question "Is this system (mind or machine) exhibiting compositionality?"

No Escape from Behavior in Evaluating Compositionality by Najoung Kim

OG hyp gen paper update! Even if you're not interested in the specific phenomenon we look at, the general framing should be of interest to a much broader audience

Kanishka Misra@kanishka.bsky.social · 4mo ago

Announcing a new version of our 2024 paper on linguistic hypothesis generation from LMs! @najoung.bsky.social and I have systematized our hypothesis generation framework, added stringent criteria for model selection, 10x-ed our learning trials, and included an epigraph from Jeff Elman 🙏!

Title page for the paper “A systematic framework for generating novel experimental hypotheses from language models”, with an epigraph from Jeff Elman describing how Rumelhart and McClelland (1982) did hypothesis generation with their connectionist network, and a figure describing our pipeline.

“All bears have a property”, “Some bears have a property”, “Bears have a property” are different in terms of how the property is generalized to a specific bear – a great example of how language constrains thought! This holds for kids, adults, and according to our new work, (V)LMs! 🧵

Title page of our paper: "Bears, all bears, and some bears. Language Constraints on Language Models' Inductive Inferences"

My lab at BU is recruiting PhD students and possibly a postdoc this year! We study humans & machines, centered around topics like meaning, generalization, evaluation methods and design, and the nature of computation and representation that underlie language and cognition. 🫴🫴

Bild

Here are the slides: docs.google.com/presentation... For context the intended audience is Lang Dev researchers who are attending a session on "what can LLMs tell us about human language". If you have any thoughts I'd love to hear them!

SLD Plenary - Najoung Kim [external share]

Whence insights? The value of delineating human and machine CogSci Najoung Kim (Boston University) Society for Language Development Annual Symposium November 6, 2025 1

docs.google.com

ever since VLMs were a thing i've been interested in how the additional visual modality changes language in meaningful ways. after negative findings after negative findings, excited to report this result! proud of our junior authors for digging into this 🐸

@yuluqin.bsky.social · last yr.

Does vision training change how language is represented and used in meaningful ways?🤔The answer is a nuanced yes! Comparing VLM-LM minimal pairs, we find that while the taxonomic organization of the lexicon is similar, VLMs are better at _deploying_ this knowledge. [1/9]

Seeing an experiment and thinking "but have they tried X? what if we do Y?" is a key part of research and a start to new discoveries. RExBench tests if coding agents can implement new extensions. It complements recent evals (eg PaperBench from OpenAI ) on replication! See 👇 for details

Sebastian Schuster@sebschu.bsky.social · last yr.

Can coding agents autonomously implement AI research extensions? We introduce RExBench, a benchmark that tests if a coding agent can implement a novel experiment based on existing research and code. Finding: Most agents we tested had a low success rate, but there is promise!

Screenshot of the RExBench preprint title page.

Repost appreciated! 🙏 ACL 2025 Ling theory & Cognitive modeling track is looking for emergency reviewers. The emergency review period is between 3/18-26, and these reviewers will be excluded from the ARR cycle. If you're interested, please sign up here! docs.google.com/forms/d/1fH7...

ACL 2025 Ling theory & Cognitive modeling track emergency reviewer volunteer form

The Linguistic Theories, Cognitive Modeling, and Psycholinguistics track at ACL 2025 is looking for emergency reviewers. The emergency reviews will take place between 18th to 26th of March, 2025. Thes...

docs.google.com