Pablo Samuel Castro

@pcastr.bsky.social

Señor swesearcher @ Google DeepMind, adjunct prof at Université de Montréal and Mila. Musician. From 🇪🇨 living in 🇨🇦. https://psc-g.github.io/

Fourth #runconference at #ICML2026 was a bit of a bust, as separate running club cancelled last minute 🫤 Still had a nice (and very humid) solo run. Tomorrow will be the last one, we'll meet at 8am at Hannam station for a run around the river, join us! naver.me/F5swAa7n

Bild
Pablo Samuel Castro@pcastr.bsky.social · 4w ago

No rain for third #runconference at #ICML2026 but it was quite humid! Tomorrow we'll join a separate running club at 8am, running around some of the main sights in Seoul, join us! Apparently anyone can join, just register here: luma.com/o0wqzdyc

Interested in the differences in strategic behavior between humans and LLMs? Come chat with me and @caroline-wang.bsky.social this Thursday afternoon, 5pm, #ICML2026 poster session in Hall A (Poster #401)!

Caroline Wang@caroline-wang.bsky.social · 6mo ago

[1/n] Just wrapped up 7 months interning with @pcastr.bsky.social at Google DeepMind and I'm so excited to share our work: arxiv.org/abs/2602.10324. TLDR: We used LLM-powered program synthesis to automatically model and discover differences between human and LLM strategic behavior

I'm excited to share our latest work “Representation Learning Enables Scalable Multitask Deep Reinforcement Learning”. In this work, we revisit a fundamental question in reinforcement learning: What if representation learning was the key ingredient behind scalable RL? 1/🧵

BildBild

I wish there were a way to increase diversity in workshop keynotes/panelists. There are a few famous researchers who end up being keynotes/panelists on multiple workshops, which means lots of other great researchers are not getting those opportunities.

Fun last #runconference at #ICLR2026, included sidewalk and beach running, a little soccer ball kicking, and going in the Ocean! also passed some lifeguards in training during my last sprint

BildBild
Pablo Samuel Castro@pcastr.bsky.social · 3mo ago

Another great #runconference this morning at #ICLR2026 ! Tomorrow is my last day so trying something different: BEACH RUN Let's meet at the pin below at 7:30am, run for a bit, then go in the water for a swim! maps.app.goo.gl/KZfuFtgsdmn4...

I will be giving a talk about this work at the "From Human Cognition to AI Reasoning: Models, Methods, and Applications" #ICLR2026 workshop 🗓️Sunday, April 26, 2026 10:30 AM BRT 📍 Room 202 C As well as a poster at 3:30pm in the same room. Come by!

Caroline Wang@caroline-wang.bsky.social · 6mo ago

[1/n] Just wrapped up 7 months interning with @pcastr.bsky.social at Google DeepMind and I'm so excited to share our work: arxiv.org/abs/2602.10324. TLDR: We used LLM-powered program synthesis to automatically model and discover differences between human and LLM strategic behavior

Applications for the global Google PhD Fellowship Program are open! Fellowships directly support grad students doing exceptional and innovative research in computer science and related fields as they pursue their PhD. Learn more and apply by April 30 at goo.gle/phdfellowship.

Google PhD fellowship program

The Google PhD Fellowship Program recognizes outstanding graduate students doing exceptional work in computer science, related disciplines, or promising research areas.

goo.gle

New paper 🚨 "Stable Deep Reinforcement Learning via Isotropic Gaussian Representations" Deep RL suffers from unstable training, representation collapse, and neuron dormancy. We show that a simple geometric insight, isotropic Gaussian representations, can fix this. Here's how 👇

BildBild

It's been such a pleasure working with @caroline-wang.bsky.social the last few months. She is a fantastic researcher, engineer, & person, and I really hope we can work together again. Check out the paper that came out of her time with us, comparing strategic behavior of humans and LLMs!👇🏽

Caroline Wang@caroline-wang.bsky.social · 6mo ago

[1/n] Just wrapped up 7 months interning with @pcastr.bsky.social at Google DeepMind and I'm so excited to share our work: arxiv.org/abs/2602.10324. TLDR: We used LLM-powered program synthesis to automatically model and discover differences between human and LLM strategic behavior