đ„Excited to share our new work: "A Practitioner's Guide to Multi-turn Agentic Reinforcement Learning"! We study what actually works for agentic multi-turn RL with varying đEnvironment, đ€Policy, and âReward. We conduct various ablations and empirical analysis on đ§©TextWorld, đ§ALFWorld, and đ§âđ»SWE-Gym.
Ruiyi Wang
@ruiyiwang.bsky.social
2nd year PhD at UCSD w/ @rajammanabrolu.bsky.social Prev: @ltiatcmu.bsky.social @umich.edu Research: Agentsđ€, Reasoningđ§ , GamesđŸ
Started a SoCal AI/ML/NLP researchers starter pack! It's a bit sparse right now, and perhaps more NLP heavy, but hey, nominate yourself and others! go.bsky.app/6QckPj9