Aly Lidayan

@aliday.bsky.social

CS PhD student at UC Berkeley studying RL and cognitive science alyd.github.io

New paper! Setting goals is so useful in complex settings because it can reduce the environment's complexity to only the goal-relevant features. But for humans, reframing comes at a cost. Could this explain why we're so bad at abandoning goals that are no longer worthwhile? 🧵

Bild

I'm presenting this 3-5:30pm on Saturday, Hall 3 #396 🌞 come chat about designing rewards and intrinsic motivation for RL + meta RL!

Aly Lidayan@aliday.bsky.social · last yr.

🚨Our new #ICLR2025 paper presents a unified framework for intrinsic motivation and reward shaping: they signal the value of the RL agent’s state🤖=external state🌎+past experience🧠. Rewards based on potentials over the learning agent’s state provably avoid reward hacking!🧵

🚨Our new #ICLR2025 paper presents a unified framework for intrinsic motivation and reward shaping: they signal the value of the RL agent’s state🤖=external state🌎+past experience🧠. Rewards based on potentials over the learning agent’s state provably avoid reward hacking!🧵

Bild