🤖🧠NEW PAPER🧠🤖 (The result of an 8-year project!) LLMs seem very different from symbolic systems. Yet LLMs excel in symbolic domains (e.g., language/code/math). How do they do it? Our finding: LLM representations have implicit symbolic structure! Link in thread ⬇️ 1/n
Jennifer Hu
@jennhu.bsky.social
Asst Prof at Johns Hopkins Cognitive Science • Director of the Group for Language and Intelligence (glint) ✨• Interested in all things language, cognition, and AI jennhu.github.io
Hello world! 👋 We’re Minds, Machines, and Brains (MMB) 👤🤖🧠 a new open access journal from @mitpress.bsky.social exploring the principles of intelligence and cognition across natural and artificial minds. Submissions open this Fall! 🔗 direct.mit.edu/mmb
Minds, Machines, and Brains | MIT Press
direct.mit.edu
The full BBS treatment from me and @futrell.bsky.social on "How linguistics learned to stop worrying and love the LMs" is now out, with all the commentaries and our response. If you "Save PDF", it will give you the whole target article + commentary + response pdf: www.cambridge.org/core/journal...
How linguistics learned to stop worrying and love the language models | Behavioral and Brain Sciences | Cambridge Core
How linguistics learned to stop worrying and love the language models - Volume 49
cambridge.org
What's more nonsensical: smashing a pumpkin using a number, or growing flowers inside a sneeze? Our paper on graded inconceivability is out now in Cognition! Come for the cognitive science 🧠🔍, stay for the whimsy 🌼🧚! 🔗Journal link: bit.ly/gradedInconCog
Now out (for realz) in Cognition: "People Make Graded Judgments About The Inconceivable" (by Hu, Sosa, & me) Free preprint: www.tomerullman.org/papers/grade... Journal link: bit.ly/gradedInconCog @jennhu.bsky.social @cognitionjournal.bsky.social
Sadly won't be at ACL in person, but check out our presentations below! 🌟 I'm giving a (remote) keynote at SCiL on 7/4! 🌟We also have a poster on probability x grammaticality, and a talk on pragmatics x Theory of Mind! Our lab is actively recruiting, so please reach out! Details at glintlab.org ✨
For a year and a half, @carorowland.bsky.social, @lehersingh.bsky.social, Marisa Casillas, Shanley Allen, and I have been meeting to discuss whether innateness is still a useful concept to think about in studying language acquisition. Here's our take: osf.io/preprints/ps...
I'm excited to share that this paper was accepted at ICLR 2026! We show that language models encode one of the most basic ingredients of a world model: the ability to distinguish plausible from implausible states. Check out the paper for more details! See you in Rio! Paper: arxiv.org/abs/2507.12553
I wrote a short article on AI Model Evaluation for the Open Encyclopedia of Cognitive Science 📕👇 Hope this is helpful for anyone who wants a super broad, beginner-friendly intro to the topic! Thanks @mcxfrank.bsky.social and @asifamajid.bsky.social for this amazing initiative!
AI Model Evaluation by Jennifer Hu: https://doi.org/10.21428/e2759450.b0757e8d
With some trepidation, I'm putting this out into the world: gershmanlab.com/textbook.html It's a textbook called Computational Foundations of Cognitive Neuroscience, which I wrote for my class. My hope is that this will be a living document, continuously improved as I get feedback.
Hopkins Cog Sci is hiring! We have two open faculty positions: one in vision, and one language. Please repost!
We are seeking candidates for two tenured/tenure-track faculty positions: One in high-level vision, written language and/or conceptual representation apply.interfolio.com/178825 One in language apply.interfolio.com/178813 Please help us spread the word!
New work to appear @ TACL! Language models (LMs) are remarkably good at generating novel well-formed sentences, leading to claims that they have mastered grammar. Yet they often assign higher probability to ungrammatical strings than to grammatical strings. How can both things be true? 🧵👇
It’s grad school application season, and I wanted to give some public advice. Caveats: -*-*-*-* > These are my opinions, based on my experiences, they are not secret tricks or guarantees > They are general guidelines, not meant to cover a host of idiosyncrasies and special cases
Interested in doing a PhD at the intersection of human and machine cognition? ✨ I'm recruiting students for Fall 2026! ✨ Topics of interest include pragmatics, metacognition, reasoning, & interpretability (in humans and AI). Check out JHU's mentoring program (due 11/15) for help with your SoP 👇
The department of Cognitive Science @jhu.edu is seeking motivated students interested in joining our interdisciplinary PhD program! Applications due 1 Dec Our PhD students also run an application mentoring program for prospective students. Mentoring requests due November 15. tinyurl.com/2nrn4jf9
New preprint! "Non-commitment in mental imagery is distinct from perceptual inattention, and supports hierarchical scene construction" (by Li, Hammond, & me) link: doi.org/10.31234/osf... -- the title's a bit of a mouthful, but the nice thing is that it's a pretty decent summary
At #COLM2025 and would love to chat all things cogsci, LMs, & interpretability 🍁🥯 I'm also recruiting! 👉 I'm presenting at two workshops (PragLM, Visions) on Fri 👉 Also check out "Language Models Fail to Introspect About Their Knowledge of Language" (presented by @siyuansong.bsky.social Tue 11-1)
Can AI models introspect? What does introspection even mean for AI? We revisit a recent proposal by Comșa & Shanahan, and provide new experiments + an alternate definition of introspection. Check out this new work w/ @siyuansong.bsky.social, @harveylederman.bsky.social, & @kmahowald.bsky.social 👇
How reliable is what an AI says about itself? The answer depends on whether models can introspect. But, if an LLM says its temperature parameter is high (and it is!)….does that mean it’s introspecting? Surprisingly tricky to pin down. Our paper: arxiv.org/abs/2508.14802 (1/n)
Due to popular demand, we are extending the CogInterp submission deadline again! 🗓️🥳 Submit by *8/27* (midnight AoE)
Excited to announce the first workshop on CogInterp: Interpreting Cognition in Deep Learning Models @ NeurIPS 2025! 📣 How can we interpret the algorithms and representations underlying complex behavior in deep learning models? 🌐 coginterp.github.io/neurips2025/ 1/4
🗓️ The submission deadline for CogInterp @ NeurIPS has officially been *extended* to 8/22 (AoE)! 👇 Looking forward to seeing your submissions!
Excited to announce the first workshop on CogInterp: Interpreting Cognition in Deep Learning Models @ NeurIPS 2025! 📣 How can we interpret the algorithms and representations underlying complex behavior in deep learning models? 🌐 coginterp.github.io/neurips2025/ 1/4
Heading to CogSci this week! ✈️ Find me giving talks on: 💬 Prod-comp asymmetry in children and LMs (Thu 7/31) 💬 How people make sense of nonsense (Sat 8/2) 📣 Also, I’m recruiting grad students + postdocs for my new lab at Hopkins! 📣 If you’re interested in language / cognition / AI, let’s chat! 😄
Excited to announce the first workshop on CogInterp: Interpreting Cognition in Deep Learning Models @ NeurIPS 2025! 📣 How can we interpret the algorithms and representations underlying complex behavior in deep learning models? 🌐 coginterp.github.io/neurips2025/ 1/4
Home
First Workshop on Interpreting Cognition in Deep Learning Models (NeurIPS 2025)
coginterp.github.io
Happy to announce the first workshop on Pragmatic Reasoning in Language Models — PragLM @ COLM 2025! 🎉 How do LLMs engage in pragmatic reasoning, and what core pragmatic capacities remain beyond their reach? 🌐 sites.google.com/berkeley.edu/praglm/ 📅 Submit by June 23rd
PragLM @ COLM '25
IMPORTANT DATES
sites.google.com
Excited to share a new preprint w/ @michael-lepori.bsky.social & Michael Franke! A dominant approach in AI/cogsci uses *outputs* from AI models (eg logprobs) to predict human behavior. But how does model *processing* (across layers in a forward pass) relate to human real-time processing? 👇 (1/12)
Check out our new work on introspection in LLMs! 🔍 TL;DR we find no evidence that LLMs have privileged access to their own knowledge. Beyond the study of LLM introspection, our findings inform an ongoing debate in linguistics research: prompting (eg grammaticality judgments) =/= prob measurement!
New preprint w/ @jennhu.bsky.social @kmahowald.bsky.social : Can LLMs introspect about their knowledge of language? Across models and domains, we did not find evidence that LLMs have privileged access to their own predictions. 🧵(1/8)
new preprint on Theory of Mind in LLMs, a topic I know a lot of people care about (I care. I'm part of people): "Re-evaluating Theory of Mind evaluation in large language models" (by Hu* @jennhu.bsky.social , Sosa, and me) link: arxiv.org/pdf/2502.21098
AI models are fascinating, impressive, and sometimes problematic. But what can they tell us about the human mind? In a new review paper, @noahdgoodman.bsky.social and I discuss how modern AI can be used for cognitive modeling: osf.io/preprints/ps...
Some things are more impossible than others. But some things might be even *more impossible* than impossible. (How) do people differentiate between the inconceivable and the merely impossible? Do language models also make similar distinctions? Check out our new preprint below!
The Red Queen believed "6 impossible things before breakfast." But what about *inconceivable* things? For your breakfast read, check out the new preprint: "Shades of Zero: Distinguishing Impossibility from Inconceivability" (by @jennhu.bsky.social , Sosa, & me) arxiv: arxiv.org/pdf/2502.20469
Hello! I'm looking to hire a post-doc, to start this Summer or Fall. It'd be great if you could share this widely with people you think might be interested. More details on the position & how to apply: bit.ly/cocodev_post... Official posting here: academicpositions.harvard.edu/postings/14723
Now hiring for two lab manager positions at Stanford! Hyo Gweon and I are coordinating joint searches since our labs collaborate frequently. Please join us! careersearch.stanford.edu/jobs/researc... and careersearch.stanford.edu/jobs/lab-coo...
(1/9) Excited to share my recent work on "Alignment reduces LM's conceptual diversity" with @tomerullman.bsky.social and @jennhu.bsky.social, to appear at #NAACL2025! 🐟 We want models that match our values...but could this hurt their diversity of thought? Preprint: arxiv.org/abs/2411.04427
LMs need linguistics! New paper, with @futrell.bsky.social, on LMs and linguistics that conveys our excitement about what the present moment means for linguistics and what linguistics can do for LMs. Paper: arxiv.org/abs/2501.17047. 🧵below.