Hayoung Jung

@hayoungjung.bsky.social

PhD student at @princetoncitp.bsky.social. Previously @uwcse.bsky.social website: hayoungjung.me

First paper of my PhD with my amazing advisors! There’s been a ton of hype and media coverage on OpenEvidence as an “AI co-pilot for clinicians”… and our long-horizon benchmark puts them to the test!! Our results suggest they are far from reliable for downstream use.

Manoel Horta Ribeiro@manoelhortaribeiro.bsky.social · 2mo ago

New preprint! We introduce a new benchmark, SciConBench, with 9.11k scientific questions derived from Cochrane Systematic Reviews. We find evidence that frontier AI agents **cannot** synthesize scientific conclusions well. A thread 🧵 w/ @hayoungjung.bsky.social & others!

🛍️Major AI companies are increasingly embedding sponsored content into chatbot conversations. Across two preregistered experiments (N=2,012), we test how effectively AI can steer consumers toward sponsored products in a realistic shopping scenario. 📝https://arxiv.org/abs/2604.04263

Bild

🚨YouTube is a key source of health info, but it’s also rife with dangerous myths on opioid use disorder (OUD), a leading cause of death in the U.S. To understand the scale of such misinformation, our #EMNLP2025 paper introduces MythTriage, a scalable system to detect OUD myth🧵

Bild

On my way to Copenhagen, where I will give an invited talk at a workshop and present this work at ICWSM! Super excited to meet everyone -- please DM me if you would like to chat!

Hayoung Jung@hayoungjung.bsky.social · 2y ago

How does YouTube’s search algorithm handle COVID🦠 misinfo in the United States🇺🇸(US) and South Africa🇿🇦(SA)? In our #icwsm '25 paper w/ @prerna6.bsky.social @tanumitra.bsky.social, we found bots in SA received significantly more misinfo in top-10 search results, which accounts for 95% of user traffic

Title and header describing