Elizabeth Mieczkowski

@emiecz.bsky.social

Studying multi-agent collaboration 🤝🧩🤖 PhD Candidate at Princeton CS with Tom Griffiths & Natalia Vélez @cocoscilab.bsky.social @velezcolab.bsky.social Prev: Cornell CS, MIT BCS

Telling two people to be selfish barely changes their behavior, unless they both know they got the same message! Our #CogSci2026 paper reveals that these group expectations only matter when we all know that we share the same expectations.

Bild

A paper casting doubt on the foundations of a bunch of nice convergence bounds in multi-agent RL (found via @quokkka.bsky.social). The technical material here is a bit out of my depth, but relevant to some stuff I do in game AI. Attempt to briefly summarize based on an initial read below. 1/n

Paradoxes of Game Theoretic Equilibria and Price of Anarchy

For decades, static solution concepts (Nash, Correlated, and Coarse Correlated Equilibria) and the Price of Anarchy (PoA) have formed the bedrock of algorithmic game theory, with no-regret learning pr...

arxiv.org

Can LLM agents coordinate in long-horizon, open-ended worlds? We test 13 LLMs in Alem, a new benchmark. Most struggle, averaging ~6% normalised return. Yet on Hard, zero-shot Gemini 3.1 Pro matches the best MARL agent after 1B training steps. Our ablations show communication matters most. 🧵

🚨New preprint and our results are rather concerning.. We find the "boiling frog" equivalent of AI use. Using large-scale RCTs, we provide *casual* evidence that AI assistance reduces persistence and hurts independent performance. And these effects emerge after just 10–15 minutes of AI use! 1/

Bild

🚨New preprint! LLM teams are being deployed at scale, yet we lack the tools to predict when they’ll succeed, fail, or how to design them. Distributed computing faced the exact same questions and figured out how to answer them. We show those insights apply directly to LLMs 🧵👇

Bild

The visual world is composed of objects, and those objects are composed of features. But do VLMs exploit this compositional structure when processing multi-object scenes? In our 🆒🆕 #ICLR2026 paper, we find they do – via emergent symbolic mechanisms for visual binding. 🧵👇

Bild

Are you a grad student who wants to give a talk at Princeton’s psychology department (in-person or on Zoom)? Nominate yourself or someone you know: forms.gle/WN2ybYMuZiW3... Priority given to non-Ivy and URM students. International applicants welcome. Deadline is this Friday (Feb 6)!

Emerging Scholars in Psychological Science Speaker Nominations

Nominate yourself or another late-stage PhD student to speak at Princeton's Department of Psychology this academic semester (Spring 2026). The Emerging Scholars in Psychological Science (ESPS) talk ...

forms.gle

I'm looking for two PhD student for Fall 2026. Both on multi-agent reinforcement learning (MARL). - Theory of MARL: experience with theory and/or MARL -Formal methods for MARL: experience with formal methods or MARL (interest in learning the other) www.khoury.northeastern.edu/programs/com...

PhD in Computer Science - Khoury College of Computer Sciences

The PhD in Computer Science program will prepare you with advanced knowledge, industry opportunities, and research experience to be a leader in the field.

khoury.northeastern.edu

Honored and excited to share that I am the winner of Nomis & Science Young Explorer Award!! Also thrilled to share that my article describing my research is out now in @science.org today! The normalization of (almost) everything www.science.org/doi/10.1126/... 1/

The normalization of (almost) everything: Our minds can get used to anything, and even crises start feeling normal

Our minds can get used to anything, and even crises start feeling normal

science.org

Honored to contribute a Journal Club piece to @natrevpsychol.nature.com! I explore Robert White's seminal "competence motivation" framework (1959) and why it remains relevant over 60 years later—from why toddlers insist on doing things themselves to designing intrinsically motivated AI. 🤖

Nature Reviews Psychology@natrevpsychol.nature.com · 10mo ago

‘Motivation reconsidered’ reconsidered Journal Club by Bella Fascendini go.nature.com/3IwYftH

Can large language models stand in for human participants? Many social scientists seem to think so, and are already using "silicon samples" in research. One problem: depending on the analytic decisions made, you can basically get these samples to show any effect you want. THREAD 🧵

The threat of analytic flexibility in using large language models to simulate human data: A call to attention

Social scientists are now using large language models to create "silicon samples" - synthetic datasets intended to stand in for human respondents, aimed at revolutionising human subjects research. How...

arxiv.org

Our new paper is out in PNAS: "Evolving general cooperation with a Bayesian theory of mind"! Humans are the ultimate cooperators. We coordinate on a scale and scope no other species (nor AI) can match. What makes this possible? 🧵 www.pnas.org/doi/10.1073/...

Evolving general cooperation with a Bayesian theory of mind | PNAS

Theories of the evolution of cooperation through reciprocity explain how unrelated self-interested individuals can accomplish more together than th...

pnas.org

So excited our paper is now out in ‪@cognitionjournal.bsky.social‬! Huge thanks to our editor and reviewers 🧠 Their thoughtful suggestions inspired Experiments 3 & 4, including a striking inverse correlation between idleness judgments and speed-up predictions

Cognition@cognitionjournal.bsky.social · last yr.

“People Evaluate Idle Collaborators Based on their Impact on Task Efficiency” 📢 New paper from: Elizabeth Mieczkowski, Cameron Rouse Turner, Natalia Vélez, & Tom Griffiths www.sciencedirect.com/science/arti... TL;DR: Sometimes it's acceptable not to help with group work 🧵👇