Atoosa Kasirzadeh

@atoosakz.bsky.social

societal impacts of AI | assistant professor of philosophy & software and societal systems at Carnegie Mellon University & AI2050 Schmidt Sciences early-career fellow | system engineering + AI + philosophy | https://kasirzadeh.org/

I’m very excited to announce the release of my first co-edited volume, "Contemporary Debates in the Ethics of Artificial Intelligence" (Wiley), alongside my wonderful co-editors @svennyholm.bsky.social and John Zerilli , on the 21 January 2026!

Sven Nyholm@svennyholm.bsky.social · 8mo ago

Out soon: *Contemporary Debates in the Ethics of Artificial Intelligence*, co-edited by me, @atoosakz.bsky.social & John Zerilli. Out on Jan 21, but preview already available on Google books: books.google.de/books?id=beK... #aiethics

🚨New paper: Reward Models (RMs) are used to align LLMs, but can they be steered toward user-specific value/style preferences? With EVALUESTEER, we find even the best RMs we tested exhibit their own value/style biases, and are unable to align with a user >25% of the time. 🧵

Bild

🚨New Paper: LLM developers aim to align models with values like helpfulness or harmlessness. But when these conflict, which values do models choose to support? We introduce ConflictScope, a fully-automated evaluation pipeline that reveals how models rank values under conflict. (📷 xkcd)

Bild

How good are the current AI agents for scientific discovery? We have answers in our new paper: lnkd.in/dxzmZXpR! We look at 4 ways AI scientists can go wrong; design experiments to show the manifestation of these failures in 2 open source AI scientists; and recommend detection strategies.

Bild

Introducing: Full-Stack Alignment 🥞 A research program dedicated to co-aligning AI systems *and* institutions with what people value. It's the most ambitious project I've ever undertaken. Here's what we're doing: 🧵

Bild

Ever since I first heard the slogan “AI as normal technology,” I’ve felt uneasy. Tonight that unease crystallised. In this 🧵 I unpack what normal hides and why the metaphor may ultimately fail to capture AI’s abnormal impacts on human life and societies. 1/n

New paper with Iason Gabriel on "Characterizing AI agents" is out! 2025 is being called the year of AI agents, with overwhelming headlines about them every day. But we lack a shared vocabulary to distinguish their fundamental properties. Our paper aims to bridge this gap. A 🧵

Bild