Nouha Dziri

@nouhadziri.bsky.social

Research Scientist at Ai2, PhD in NLP 🤖 UofA. Ex GoogleDeepMind, MSFTResearch, MilaQuebec https://nouhadziri.github.io/

Had the great pleasure this week visiting @mbzuai for the trustworthy foundation models symposium where I spoke alongside the brilliant and humble Tomas Mikolov 😍 Tomas is a true pioneer who invented word2vec, RNNLLM toolkit and other works without which Transformers might not have emerged this fast

Bild

SUPER thrilled that our #NAACL2025 paper got the runnerup BEST paper award 😍😍🏆🏆🏆🚀🚀 We show that people rely 30% more on LLMs when they use emphatic expressions (eg "Sure, happy to help") even though the answer is wrong and 10% more when the task involves math questions 😵 📜 arxiv.org/pdf/2407.07950

arxiv.org

Kaitlyn Zhou@kaitlynzhou.bsky.social · last yr.

Thrilled that our paper won 🏆 Best Paper Runner-Up 🏆 at #NAACL25!! Our work (REL-A.I.) introduces an evaluation framework that measures human reliance on LLMs and reveals how contextual features like anthropomorphism, subject, and user history can significantly influence user reliance behaviors.

Super excited to be speaking at the #icml2025 on Computer Use Agents this summer in Vancouver 🇨🇦 alongside a stellar lineup of speakers! ⏰Submit your work by *May 18, 2025* 📄https://icml-computeruseagents.com

Bild
ComputerUseAgents Workshop@workshopcua.bsky.social · last yr.

Join us, and submit your best work @ www.icml-computeruseagents.com Incredible lineup of speakers at #CUA #WORKSHOP @ #ICML2025! Speakers part 1/3: Nouha Dziri (AI2) @nouhadziri.bsky.social Ruslan Salakhutdinov (CMU, Meta) @rsalakhu.bsky.social Sercan Arik (Google Cloud)

I still can't comprehend how an AC (a professor) accepts the role but then never responds back to emails and never completes their tasks😖! We need urgent 5 emergency reviewers to complete reviews for ACL by the end of today. Area: Ethics, Bias, and Fairness. Please reach out if you can help! Thanks🙏

And that was a wrap #NeurIPS2024 was intense, fast-paced, rich, packed🔥 Super happy with the success of Sys2 Reasoning: a true concentration of top AI figures who pioneered the field Yoshua Benjio@yoshuabengio.bsky.social Dima Bahdaneau, Melanie Mitchel, Jason Weston, Dawn Song, Joshua Tenenbaum👇

Bild

#NeurIPS2024 remains my favorite venue, excited to be there next week🥳Reach out to chat, I will *Speak at the panel of the Inference-Time Algorithms tutorial ⏰Dec 10 *Give a talk at the Language Gamification workshop: in-context learning /scaling inference ⏰Dec 14 More👇

Enjoying being at Mountain View today and feeling like summer with Christmassy vibe 18C🏖️😎 Had the pleasure to speak at the MLCommons panel for the release of their first safety benchmark AlLuminate v1.0. It’s incredible to witness such great open community work coming together🎉 Use AlLuminate!

BildBildBild

🚀New work🔥 CREATIVITY Index 🔥 This work is SO close to my heart, I loved every part of the experiments. It provided me with so much scientific fulfillment. Intellectual works have become so rare in the hysterical race of AI, so if you care about science give this work a read! Read below hot takes:

Bild
Ximing Lu@gximing.bsky.social · 2y ago

Are LLMs 🤖 as creative as humans 👩‍🎓? Not quite! Introducing CREATIVITY INDEX: a metric that quantifies the linguistic creativity of a text by reconstructing it from existing text snippets on the web. Spoiler: professional human writers like Hemingway are still far more creative than LLMs! 😲

Ok ✨an inaugural post✨ l had the pleasure yesterday to lead the agentic AI safety sessions in Columbia/Mozilla safety convening and meet with the 🇫🇷French Minister of AI, Madame Clara Chappaz with whom I discussed the urgency of regulatory frameworks to enable safe deployment of agents👇

Bild