New paper! Setting goals is so useful in complex settings because it can reduce the environment's complexity to only the goal-relevant features. But for humans, reframing comes at a cost. Could this explain why we're so bad at abandoning goals that are no longer worthwhile? 🧵
Aly Lidayan
@aliday.bsky.social
CS PhD student at UC Berkeley studying RL and cognitive science alyd.github.io
I'm presenting this 3-5:30pm on Saturday, Hall 3 #396 🌞 come chat about designing rewards and intrinsic motivation for RL + meta RL!
🚨Our new #ICLR2025 paper presents a unified framework for intrinsic motivation and reward shaping: they signal the value of the RL agent’s state🤖=external state🌎+past experience🧠. Rewards based on potentials over the learning agent’s state provably avoid reward hacking!🧵
1/3 Out now: new paper on people's perception of AI (robot) creativity! Core finding: we attribute more creativity to a creative act if people not only see the final artwork, but also its creation process & the robot making it. Video: vimeo.com/1073134853 Open-access paper: doi.org/10.1145/3711...
🚨Our new #ICLR2025 paper presents a unified framework for intrinsic motivation and reward shaping: they signal the value of the RL agent’s state🤖=external state🌎+past experience🧠. Rewards based on potentials over the learning agent’s state provably avoid reward hacking!🧵