Steve Rathje

@steverathje.bsky.social

Incoming Assistant Professor of HCI at Carnegie Mellon studying the psychology of technology. NSF postdoc at NYU, PhD from Cambridge, BA from Stanford. stevenrathje.com

🚨 New preprint 🚨 We developed a sycophancy taxonomy based on prior literature and surveyed 106 experts. 94% agreed it's a serious problem. But they substantially disagreed about which behaviors actually count as sycophancy.

Bild

People prefer to use “sycophantic” AI systems that reinforce their pre-existing beliefs. In a new paper (n=7,227), we found that people enjoyed interacting with sycophantic AI chatbots more than interacting with neutral chatbots or “disagreeable” chatbots that challenged their beliefs.

Bild

Recent preprint with Oriel FeldmanHall and Matt Nassar, showing clear neural evidence that adolescent (13-15 yrs) social media and smartphones use are associated with blunted reward signaling in the ventral striatum, the brain's reward processing hub, and worse mental health. osf.io/preprints/ps...

Bild

This shapes up to be a deep dive into the psychological effects of sycophant AI, making you overconfident and blind to your own biases. Would be interesting to see a comparison to social media effects on the same. After all, tribes are sycophant networks too.

Steve Rathje@steverathje.bsky.social · 3mo ago

We’ve updated our pre-print on the effects of AI sycophancy with four additional studies (total n = 7,227). Here is a brief summary of our new findings (🧵1/n):

Screenshot of abstract: AI can be a powerful tool for opening people up to new perspectives, yet people may prefer to use “sycophantic” (or overly agreeable and validating) AI systems that reinforce their pre-existing beliefs. Across seven studies (total n = 7,227), we found that people enjoyed interacting with sycophantic AI chatbots more than interacting with neutral chatbots or “disagreeable” chatbots that challenged their beliefs. Brief conversations with sycophantic chatbots about political or personal topics increased attitude extremity and certainty, with most effects persisting for at least one week. Sycophantic chatbots also inflated people’s perceptions that they were better than average on desirable traits (e.g., intelligence, empathy). Moreover, people who interacted with sycophantic (rather than disagreeable) AI bet more money that they scored better than average on tasks measuring these traits (approximately 6 cents more out of 75 possible cents), demonstrating that sycophancy can affect costly decisions. Participants consistently rated sycophantic chatbots as more “unbiased” than disagreeable chatbots, even though third-party raters viewed these chatbots as equally biased, suggesting that people may be blind to biases in AI output that aligns with their views. People were more receptive to chatbots that presented opposing information when that information was presented in a validating way, and individuals who scored higher on a measure of intellectual humility were also more receptive to disagreeing chatbots. Altogether, these results suggest that people’s preference for, and blindness to, sycophantic AI risks creating AI “echo chambers” that increase attitude extremity and lead to overconfident beliefs and decisions.

AI is forcing us to reconcile our own psychology, & the subtle/unnoticed ways in which we can be manipulated, as we never have before. It's nothing new with AI, it's just a new discussion for the public in general.

Steve Rathje@steverathje.bsky.social · 3mo ago

Nevertheless, people enjoyed sycophantic chatbots more than disagreeable ones, chose to interact with them more, viewed them as warmer and more competent, and felt a stronger "sense of connection" with them.

In my latest podcast episode, I discuss the psychology of virality with @steverathje.bsky.social, explore how agreeable AI chatbots may influence our beliefs, and examine how scientists can communicate effectively in a noisy, polarized media environment. matthewfacciani.substack.com/p/the-psycho...

The Psychology of Virality in the Age of AI

Steve Rathje joins me to discuss why conflict goes viral, how social media shapes polarization, and what AI chatbots mean for belief and bias.

matthewfacciani.substack.com

Enjoyed talking with @sudkrc.bsky.social on one of my favorite podcasts, the Stanford Psychology Podcast! We discuss how I got into psychology (it all began at Stanford), my recent work on the psychology of virality and sycophantic AI, and much more.

Stanford Psychology Podcast@stanfordpsypod.bsky.social · 8mo ago

NEW EPISODE OUT🗣️!! In this episode, Su @sudkrc.bsky.social chats with Dr. Steve Rathje @steverathje.bsky.social on why certain content spreads rapidly online and offline! LISTEN NOW🎧: open.spotify.com/episode/7CoK...

While studies find that moral outrage & negativity goes viral on social media, this is also true of the offline world. Gossip is also mostly negative & about people we dislike I explain why some ideas go viral--but most don't with @steverathje.bsky.social www.powerofusnewsletter.com/p/why-some-i...

Why Some Ideas Go Viral—and Most Don’t

What decades of research reveal about why certain content spreads—and how social forces shape what we all see.

powerofusnewsletter.com

🚨 New working paper 🚨 Can LLMs with reasoning + web search reliably fact-check political claims? We evaluated 15 models from OpenAI, Google, Meta, and DeepSeek on 6,000+ PolitiFact claims (2007–2024). Short answer: Not reliably—unless you give them curated evidence. arxiv.org/abs/2511.18749

BildBild

🚨 New preprint 🚨 Across 3 experiments (n = 3,285), we found that interacting with sycophantic (or overly agreeable) AI chatbots entrenched attitudes and led to inflated self-perceptions. Yet, people preferred sycophantic chatbots and viewed them as unbiased! osf.io/preprints/ps... Thread 🧵

Abstract and results summary