Jessica Hullman

@jessicahullman.bsky.social

Ginni Rometty Prof @NorthwesternCS | Fellow @NU_IPR | AI, decisions metascience | Blog @statmodeling substack.com/@jessicahullman | Direct hullmanlab.northwestern.edu

AI impact evaluation is still often very rudimentary in practice. I was glad to be part of this preprint led by @rbly.bsky.social on how to think about the right evidence standards for evaluating models in clinical settings.

Robin Blythe@rbly.bsky.social · 3w ago

New #healtheconomics #preprint in which we propose a proportionality argument for evaluation of clinical AI: that if the model leads to changes in the way patients are treated, then it must meet the evidentiary standards of that treatment.

Nihar Shah did a heroic experiment for TMLR: he spent 20-25 hours over two weeks interviewing authors of seemingly low-quality submissions about their own papers. He confirmed what we all suspected: people submitting these papers have *no idea* what is going on in them.

Transactions on Machine Learning Research@tmlrorg.bsky.social · 4w ago

TMLR has faced a deluge of submissions, necessitating stricter desk rejection policies due to limited reviewer capacity Co-EiC Nihar Shah reached out to authors of 10 papers slated for desk reject. Could they answer questions about their *own* submission? medium.com/@TmlrOrg/ask...

I'm getting asked to review a LOT of heavily AI-written sloppy papers. Very often from top general science journals. Most have policies to ensure author accountability. They don't seem to be working. Reviewers can incentivize better policy by temporarily creating friction. 1/

My new go-to review request response, sadly: Hi, I declined your request, as the abstract is flagged 100% AI-generated by Pangram. Deciphering what claims authors intended vs originated from AI is not a good use of my time. Jessica Also have one to reply to prospective students

If you feel the moral stakes of this moment are “universities must be defended,” this may not be a reassuring post. It backs up to ask “why,” and finds the answer not self-evident. But strangely, honest wrestling with fundamentals reassures me more in the end than polemic.

Jessica Hullman@jessicahullman.bsky.social · 2mo ago

The uncertain future of academia got me interested in the history of US science policy. It's surprising how fragile the basic vs applied research distinction behind the postwar “social contract for science” is. There are takeaways for what it means to defend universities now🧵

About a year ago, I wrote skeptically about LLMs in peer review -- not because of skepticism about their inherent capabilities, but because I don't want the research community to optimize for the taste of any one person/system. What's changed since then?

What’s next for machine learning peer review?

A bit over a year ago, I wrote about the dangers of using LLMs for peer review. The most serious concern I had was algorithmic monoculture: the research community would collectively end up optimizing ...

bryanwilder.substack.com

Multiverse Analyses “[Only 6/152 (3.9%)] studies discussed whether their competing specifications were defensible or principled (distinguishing between equivalent, non-equivalent, or uncertain specifications) in the sense of Del Giudice & Gangestad (2021) [screenshot below]” doi.org/10.1177/2515...

Bild
Alejandro Sandoval-Lentisco@asandovall.bsky.social · 3mo ago

How often are multiverse analyses implemented, not just cited/discussed? How are they implemented? 🧵 New preprint on the uptake and implementation of multiverse-style analyses 👇 www.biorxiv.org/content/10.6...

Newly proposed rules from the federal government will irreparably damage US science: please express your dissent and post a comment before the OMB public comment period closes this Monday, July 13. Read about the endgame here: www.theguardian.com/commentisfre...

Is the US trying to make scientists’ work so difficult that they simply give up? | Daniel Malinsky

New Trump administration rules would undermine longstanding research practices. It’s death by a thousand cuts

theguardian.com