Sign in

Anthony GX-Chen

@agx-chen.bsky.social
26 followers 21 following 7 posts

PhD student at NYU CILVR. Prev: Master's at McGill / Mila. || RL, ML, Neuroscience. im-ant.github.io

PostsRepliesMedia
Anthony GX-Chen @agx-chen.bsky.social · 05/10/2025
This work has been accepted to #COLM2025. If you are in Montreal this week for COLM and would like to chat about this (or anything related to discovery / exploration / RL), drop me a note! Poster session 2: Tuesday Oct 7, 4:30-6:30pm Poster number 68
011
Anthony GX-Chen @agx-chen.bsky.social · 16/05/2025
Language model (LM) agents are all the rage now—but they may exhibit cognitive biases when inferring causal relationships! We evaluate LMs on a cognitive task to find: - LMs struggle with certain simple causal relationships - They show biases similar to human adults (but not children) 🧵⬇️
Example of the Blicket Test experiment. A subset of objects activate the machine following an unobserved rule ("disjunctive" / "conjunctive"). The agent needs to interact with the environment by placing objects on/off the machine to figure out the rule.
171
Reposted by Anthony GX-Chen
Alison Gopnik @alisongopnik.bsky.social · 15/05/2025
Fascinating preprint from with our "blicket detector" paradigm from Chen et al at NYU& Mila. LLM's make the same causal inference mistakes that adults make but 4 year olds don't! Of course, models are trained on adult data, kids figure it out for themselves. im-ant.github.io/publications...
im-ant.github.io
2182