Anthony GX-Chen

@agx-chen.bsky.social

PhD student at NYU CILVR. Prev: Master's at McGill / Mila. || RL, ML, Neuroscience. https://im-ant.github.io/

Language model (LM) agents are all the rage now—but they may exhibit cognitive biases when inferring causal relationships! We evaluate LMs on a cognitive task to find: - LMs struggle with certain simple causal relationships - They show biases similar to human adults (but not children) 🧵⬇️

Example of the Blicket Test experiment. A subset of objects activate the machine following an unobserved rule ("disjunctive" / "conjunctive"). The agent needs to interact with the environment by placing objects on/off the machine to figure out the rule.