Kshitish Ghate

@kghate.bsky.social

PhD student @ UWCSE; MLT @ CMU-LTI; Responsible AI https://kshitishghate.github.io/

Happy to share that I’m presenting 3 research projects at AIES 2025 🎉 1️⃣Gender bias over-representation in AI bias research 👫 2️⃣Stable Diffusion's skin tone bias 🧑🏻🧑🏽🧑🏿 3️⃣Limitations of human oversight in AI hiring 👤🤖 Let's chat if you’re at AIES or read below/reach out for details! #AIES25 #AcademicSky

🚨New paper: Reward Models (RMs) are used to align LLMs, but can they be steered toward user-specific value/style preferences? With EVALUESTEER, we find even the best RMs we tested exhibit their own value/style biases, and are unable to align with a user >25% of the time. 🧵

Bild

🚨New Paper: LLM developers aim to align models with values like helpfulness or harmlessness. But when these conflict, which values do models choose to support? We introduce ConflictScope, a fully-automated evaluation pipeline that reveals how models rank values under conflict. (📷 xkcd)

Bild

Excited to announce our #NAACL2025 Oral paper! 🎉✨ We carried out the largest systematic study so far to map the links between upstream choices, intrinsic bias, and downstream zero-shot performance across 131 CLIP Vision-language encoders, 26 datasets, and 55 architectures!

Bild

Excited to announce our #NAACL2025 Oral paper! 🎉✨ We carried out the largest systematic study so far to map the links between upstream choices, intrinsic bias, and downstream zero-shot performance across 131 CLIP Vision-language encoders, 26 datasets, and 55 architectures!

Bild