Did you know that LLMs suffer from serious mode collapse? For example, if you ask models to tell you a joke, they almost always tell you the same joke? This is true across samples and even across model families! Why does this happen? Can we improve it?
I'm recruiting 1-2 PhD students to work with me at the University of Colorado Boulder! Looking for creative students with interests in #NLP and #CulturalAnalytics. Boulder is a lovely college town 30 minutes from Denver and 1 hour from Rocky Mountain National Park 😎 Apply by December 15th!
✨I am on the faculty job market in the 2024-2025 cycle!✨ My research centers on advancing Responsible AI, specifically enhancing factuality, robustness, and transparency in AI systems. If you have relevant positions, let me know! lasharavichander.github.io Please share/RT!
Abhilasha Ravichander - Home
lasharavichander.github.io
Why and when do preference annotators disagree? And how do reward models + LLM-as-Judge evaluators handle disagreements? Michael explored these questions in a new ✨preprint✨ from his @ai2.bsky.social internship with me!
✨EMNLP Paper! ✨ Have you ever constructed a table to organize your literature review process? Can we use LMs to generate these automatically? We are excited to present ArxivDIGESTables 🍽️ a study of collecting, generating, and evaluating 🎓 scientific literature review tables 📃!