Josh McDermott

@joshhmcdermott.bsky.social

Working to understand how humans and machines hear. Prof at MIT; director of Lab for Computational Audition. https://mcdermottlab.mit.edu/

If at CCN, please check out three posters from our lab today. In the morning session: A62 - Deep Learning Models of Attention Reveal Peripheral Encoding as a Bottleneck for Selective Listening Through Cochlear Implants, by Annesya Banerjee

SANE 2026 will be held on Friday October 30, 2026 at MIT, in Cambridge, MA. Confirmed speakers so far include Berrak Sisman (JHU) @berraksisman.bsky.social, Bryan Pardo (Northwestern), Henry Li (Google), and Ruohan Gao (UMD). More details + registration instructions at www.saneworkshop.org/sane2026/

SANE 2026 - Speech and Audio in the Northeast

SANE is a series of workshops gathering researchers and students in speech and audio from the Northeast of the American continent.

saneworkshop.org

1. How common is LLM use in scientific publishing, and how does it vary across field, publisher, journal prestige, author demographics etc.? @kylesiler.bsky.social has new paper in PNAS that addresses this question on a massive scale: 7.3 million papers from Elsevier, PLOS, MDPI, and Frontiers.

The diffusion of large language models in published academic articles | PNAS

Large language models (LLMs) are rapidly changing academic research, raising questions of who is adopting these tools and under what conditions. Th...

pnas.org

Russell Vought, Director of the OMB, has issued a set of proposed changes that would dramatically alter federal grant funding. However, we can each take action to prevent these from taking effect. Here's how, a 🧵 🧪 1/n www.science.org/content/arti...

White House seeks to tighten political oversight of grantmaking

Sweeping proposed rule, now open for comments, would also restrict foreign collaborations and remove federal funding for open-access fees

science.org

🪼 Your tax dollars paid a scientist to squeeze jellyfish through cheesecloth. Osamu Shimomura processed tens of thousands on an NSF grant to understand why they glow. He found a protein that glowed green under UV light. Called it "green protein." Nobody cared. For decades.

Bild

Two more posters from our lab at ARO today: T152 A Model of Speech Recognition Reproduces Signatures of Human Speech Perception and Reveals Mechanisms of Contextual Integration, by Gasser Elbanna T171 Hearing-Impaired Deep Neural Networks Predict Real-World Hearing Difficulties, by Mark Saddler

If at ARO today, check out the presentations from our lab: talk at 3:15pm: Optimized Models of Uncertainty Explain Human Confidence in Auditory Perception, by Lakshmi Govindarajan poster 28: In-Silico fMRI Experiments Enable Comparisons of Speech Models to Human Auditory Cortex, by Gasser Elbanna

If at ARO, please check out four posters from our lab today: SA53 A Deep Learning Framework for Understanding Cochlear Implants, by Annesya Banerjee SA189 Robustness to Noise Reveals Cross-Culturally Consistent Properties of Pitch Perception for Harmonic and Inharmonic Sounds, by Malinda McPherson

If you are at NeurIPS I encourage you to check out Gasser's poster showing his ongoing work on models of speech perception.

Gasser Elbanna@gelbanna.bsky.social · 8mo ago

If you’re still at #NeurIPS2025, come say hi at our poster at @unireps.bsky.social in Ballroom 20D! I'm presenting work co-led by Ivy Brundege and me, with @joshhmcdermott.bsky.social, showing that in silico fMRI experiments of speech models reveal notable discrepancies with human auditory cortex.

If you are at APAN today, check out these posters from members of our lab: 1. Cross-culturally shared sensitivity to harmonic structure underlies some aspects of pitch perception - Malinda McPherson-McNato

Want to make publication-ready figures come straight from Python without having to do any manual editing? Are you fed up with axes labels being unreadable during your presentations? Follow this short tutorial including code examples! 👇🧵

Bild