Gautier Hamon
@hamongautier.bsky.social
PhD student at INRIA Flowers team. MVA master reytuag.github.io/gautier-hamon/
🚀 Introducing 🧭MAGELLAN—our new metacognitive framework for LLM agents! It predicts its own learning progress (LP) in vast natural language goal spaces, enabling efficient exploration of complex domains.🌍✨Learn more: 🔗 arxiv.org/abs/2502.07709 #OpenEndedLearning #LLM #RL
MAGELLAN: Metacognitive predictions of learning progress guide...
Open-ended learning agents must efficiently prioritize goals in vast possibility spaces, focusing on those that maximize learning progress (LP). When such autotelic exploration is achieved by LLM...
arxiv.org
we are recruiting interns for a few projects with @pyoudeyer in bordeaux > studying llm-mediated cultural evolution with @nisioti_eleni @Jeremy__Perez > balancing exploration and exploitation with autotelic rl with @ClementRomac details and links in 🧵 please share!
1/⚡️Looking for a fast and simple Transformer baseline for your RL environment in JAX ? Sharing my implementation of transformerXL-PPO: github.com/Reytuag/tran... The implementation is the first to attain the 3rd floor and obtain advanced achievements in the challenging Craftax
Now that @jeffclune.bsky.social and @joelbot3000.bsky.social are here, time for an Open-Endedness starter pack. go.bsky.app/MdVxrtD
🚨New preprint🚨 When testing LLMs with questions, how can we know they did not see the answer in their training? In this new paper we propose a simple out of the box and fast method to spot contamination on short texts with @stepalminteri.bsky.social and Pierre-Yves Oudeyer !