At NeurIPS today through Sunday! Today I'll be presenting our spotlight paper on evaluating LLM world models at the 4:30pm poster session (#2301). On Saturday I'll be co-organizing the Behavioral ML workshop. Hope to see you there! Paper: arxiv.org/abs/2406.03689 Workshop: behavioralml.org
Evaluating the World Model Implicit in a Generative Model
Recent work suggests that large language models may implicitly learn world models. How should we assess this possibility? We formalize this question for the case where the underlying reality is govern...
arxiv.org