James Zou

@jameszou.bsky.social

@Stanford Professor. AI for science and medicine.

Excited to share our new papers at #ICLR2026 on (multi-)agents, efficient reasoning, long context, better tokenizers and scientific applications 🚀 My awesome students and collaborators will be presenting them at the main conference this week; check it out! 👇

Bild

As speech models are being deployed in real-world taxi and emergency service settings, the failure to accurately transcribe named entities can cause delays and errors in critical settings.

What do LLMs think about on their own, when we let them think freely? We generated 250K “daydream” samples across models 🧠 GPT → coding Qwen → multiple-choice math exams Llama → literature DeepSeek → math, religion, psychology

Next week @jameszou.bsky.social & colleagues will host a conference where all the papers are written by AI agents & reviewed by them too. What do you reckon? A good chance to put AIs through their paces? Or a way to divert AI slop from elsewhere? 🧪🤖 My story here: www.nature.com/articles/d41...

AI bots wrote and reviewed all papers at this conference

Event will assess how reviews by models compare with those written by humans.

nature.com

"We found a troubling emergent behavior in LLM. —When LLMs compete for social media likes, they start making things up. —When they compete for votes, they turn inflammatory/populist. —When optimized for audiences, LLMs inadvertently become misaligned." → Moloch's Bargain @jameszou.bsky.social #AI

Bild