Matt Groh

@mattgroh.bsky.social

Assistant professor at Northwestern Kellogg | human AI collaboration | computational social science | affective computing

Can we train people to better detect deepfakes and not be overly skeptical of everything?🤔 We ran a 30-minute training for 🕵️ intelligence analysts with an embedded randomized experiment The short answer is yes; the key is learning which kind of artifacts point to AI 🧵

BildBildBild

AI can help us humans better understand how we connect. Empathy is something most people feel strongly but fail to communicate effectively. Just see how people respond to someone passed over for promotion Insights from 3k convos btw 968 people and LLMs from our new preprint 🧵

BildBild

🚨 New in @natmachintell.nature.com 🚨 We collected 9000+ annotations of empathic communication in convos from experts, crowds & LLMs across 4 NLP/comms/psych frameworks LLM judgment exceeds crowds' reliability & nearly matches experts Soft skills can now be reliably measured by LLMs 🧵

Bild

Happy 2026!!! In the spirit of reflecting on the past year as we begin a new one, I am sharing what our lab has been up to. Read on to learn about the q's we’re asking, find links to papers and media appearances, and see a pic of the lab at a Cubs game!

BildBild

When are LLMs-as-judge reliable? That's a big question for frontier labs and it's a big question for computational social science. Excited to share our findings (led by @aakriti1kumar.bsky.social!) on how to address this question for any subjective task & specifically for empathic communications

@aakriti1kumar.bsky.social · last yr.

How do we reliably judge if AI companions are performing well on subjective, context-dependent, and deeply human tasks? 🤖 Excited to share the first paper from my postdoc (!!) investigating when LLMs are reliable judges - with empathic communication as a case study 🧐 🧵👇

💡New paper at #CHI2025 💡 Large scale experiment with 750k obs addressing (1) How photorealistic are today's AI-generated images? (2) What features of images influence people's ability to distinguish real/fake? (3) How should we categorize artifacts?

Bild

📣 📣 Postdoc Opportunity at Northwestern Dashun Wang and I are seeking a creative, technical, interdisciplinary researcher for a joint postdoc fellowship between our labs. If you're passionate about Human-AI Collaboration and Science of Science, this may be for you! 🚀 Please share widely!

Bild

V2 of the Human and Machine Intelligence 😊🤖🧠 is in the books! So many fantastic discussions as we witnessed the frontier of AI shift even further into hyperdrive✨ Props to students for all the hard work and big thanks to teaching assistants and guest speakers 🙏

BildBild

2024 marks the official launch of the Human-AI Collaboration Lab, so I wrote a one page letter to introduce the lab, share highlights, and begin a lab tradition of reflecting on the year and sharing what we're working on in an easy to digest annual letter to share with friends and colleagues.

Bild

I'm teaching my second iteration of "Human and Machine Intelligence" 🧠🤖 for Kellogg MBAs. I updated the syllabus with a couple 2024 books + new lectures and readings. What else do you think MBA students should be reading on this topic? docs.google.com/document/d/1...

Copy of W25_MORS950_Human_and_Machine_Intelligence.docx

Human and Machine Intelligence (MORS 950) Professor Matt Groh | he/him/his | matthew.groh@kellogg.northwestern.edu Section 31 | Evanston Section 81 | Chicago Office Hours: Schedule available on Ca...

docs.google.com

New paper out in @ScienceMagazine! In 8 studies (multiple platforms, methods, time periods) we find: misinformation evokes more outrage than trustworthy news, when it does it's shared more + ppl are less likely to read before sharing. w/ @killianmcl1 @Klonick @mollycrockett 🧵👇

Bild

The ability for thoughtful people to spot AI-generated poetry in a couple seconds vs. the study's participants reveals classic problems inherent to Imitation Game research: Lack of domain expertise & lack of knowledge of AI's capabilities and limitations -> falling for & even preferring simulacra

Bild
Jessica Hullman@jessicahullman.bsky.social · 2y ago

So easy. 10 out of 10 and pretty sure I spent no more than 1.5 seconds per item. As someone who was once cared a lot about poetry, most of the poetry I see that's written by LLMs seems undeserving of even being called poetry.

I'm recruiting a PhD student to join the Human AI Collaboration lab at Kellogg, NU CS, and @nicoatnu.bsky.social If you're excited about computational social science, LLMs, digital experiments, real-world problem solving, this could be a great fit Please reshare! Deets 👇

Bild

I created a starter pack for human-AI collaboration in case that this would be helpful for some newcomers, which overlaps with several existing starter packs. Let me know if you would like to be added. go.bsky.app/MRV1Wa1

Post nicht verfügbar.

Why does ChatGPT outperform physicians + ChatGPT in this clinical vignette study? User error seems to be the culprit If users don't know how to interact with the technology (however easy it may seem), then the experiment misses out on what would happen if participants had basic knowledge LLMs

BildBild
Eric Topol@erictopol.bsky.social · 2y ago

There are now 5 reports like this—#AI performing better than physicians + AI—and we don’t have the explanation for why yet (hybrid was supposed to be best) Gift link nytimes.com/2024/11/17/hea…