Lujain Ibrahim

@lujain.bsky.social

not really on here

🚨 New preprint 🚨 We developed a sycophancy taxonomy based on prior literature and surveyed 106 experts. 94% agreed it's a serious problem. But they substantially disagreed about which behaviors actually count as sycophancy.

Bild

New! Friendlier chatbots make more mistakes. Latest Oxford research published in Nature tested 5 AI models and 400,000+ responses. Warm models made 10–30% more factual errors and were 40% more likely to agree with users' false beliefs, even on medical advice and conspiracy theories. 1/2

BildBild

At FAccT today? Hear from @oii.ox.ac.uk DPhil student @lujain.bsky.social presenting her co-authored research paper ‘Promising Topics for U.S.–China Dialogues on AI Risks and Governance’ in the AI Regulation session, 11.09am today. #FAccT2025 Read the paper: dl.acm.org/doi/10.1145/...

Promising Topics for US–China Dialogues on AI Risks and Governance | Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency

dl.acm.org

Dear ChatGPT, Am I the Asshole? While Reddit users might say yes, your favorite LLM probably won’t. We present Social Sycophancy: a new way to understand and measure sycophancy as how LLMs overly preserve users' self-image.

Bild

📈Out today in @PNASNews!📈 In a large pre-registered experiment (n=25,982), we find evidence that scaling the size of LLMs yields sharply diminishing persuasive returns for static political messages.  🧵:

BildBild

🚨 I'm recruiting 2x postdocs and 1-2 DPhil (PhD) students at Oxford to work on AI, Privacy-Enhancing Technologies, and public interest technology research. Interested in human-centred and critical approaches to study the impact of data and algorithms on society? Join us next year!

Recruiting 2x three-year postdoctoral researchers and 1-2 PhD students – Synthetic Society

The Synthetic Society research team at the Oxford Internet Institute invites applications from enthusiastic and motivated candidates for two postdoctoral positions and 1-2 PhD positions in 2025, worki...

syntheticsociety.oii.ox.ac.uk

want to hear Rob Gorwa and I talk about what content moderation of uploaded AI models on platforms like Hugging Face, GitHub and Civitai tells us about the future of who analyses open source dual use models? Mar 6 1230ET online organised by Berkman Klein. register: cyber.harvard.edu/events/moder...

Moderating Model Marketplaces

RSM welcomes Robert Gorwa and Michael Veale for a discussion of their research into the governance questions raised by the moderation of model marketplaces.

cyber.harvard.edu

designers of bluesky! we're hiring a short-term designer for a mozilla-funded simple educational game(ish) / interactive explainer on social media algorithmic feeds. email me aae322@nyu.edu for more info!

I remember a friend of mine searching for how best to describe a felt sense of "cultural kinship." When you feel close to a culture that isn't yours; when you feel a connection that's warm, pleasant, a bit defensive even. A beautiful feeling I have experienced in the past but struggle to describe