Early results, but! I’ve scraped 100,000+ law review articles (maybe~2x the next-largest study) and am using language models to classify their footnotes. Having fun seeing how tech law scholarship shakes out. So far, STS is overrepresented; sociology is slightly under.
dan bateyko
@dbateyko.bsky.social
maybe the hard stuff's inside, hidden — like bones, as opposed to an exoskeleton. @CornellInfoSci https://dbateyko.info
To make govt services more accessible & accountable, @baricks.bsky.social, @boston.gov’s aleja jimenez jaramillo & @megyoung0.bsky.social call on state & local govts to explore collectively-governed “in-house” alternatives to commercial language translation tech. datasociety.net/research-lib...
The 16th Edition of my Internet Law casebook is out! Overview: internetcasebook.com Table of contents: www.semaphorepress.com/downloads/In... PDF download (suggested $30): semaphorepress.com/InternetLaw_... Print-on-demand ($75.10): www.amazon.com/dp/1943689245 A brief thread on the update:
Just arrived at ICML 🇰🇷😍 Get up early tomorrow to hear me talk about how (not) to solve the peer review crisis, or find me at one of my poster presentations. Paper links: ✅ AI Peer Review: arxiv.org/abs/2605.03202 ✅ SWE-chat: arxiv.org/pdf/2604.20779
Can you boost your AI review scores by asking an LLM to rewrite your paper? Yes! We call it paper laundering Our @icmlconf.bsky.social spotlight paper argues current AI reviewers aren't ready to automate peer review, and outlines what a science of peer review automation should look like 🧵👇 #ICML2026
Are you at ICML next week? Feel like your decision-making for which sessions to attend might not be risk minimizing? Don't incur (swap) regret and come to my, @aaroth.bsky.social, and @ncollina.bsky.social's tutorial Monday on multicalibration, decision-making, and collaborative learning!
it's the last day of #FAccT2026 in Montreal (bonjour hiii), and I'm presenting my paper with Diag Davenport on the promise of #PublicInterestTech clinics for training the next gen of critical sociotechnical thinkers. 🏆 plus we got an honorable mention!? come thru @ 10:45! 📄 tiny.cc/pit-clinics-26
I'm so excited to attend #FAccT2026 in 🇨🇦Montreal🇨🇦 to present "Tradeoffs are Domain Dependent: Improving Accuracy and Fairness in Property Tax Assessments" by Evelyn Smith, me, Chris Berry, @jacobsgoldin.bsky.social, and Dan Ho!! 🔗: arxiv.org/pdf/2605.15020
Excited to attend FAccT 2026 in Montreal this week! Let me know if you'll be there and want to catch up :) I'll be presenting a paper with @allisonkoe.bsky.social about ad delivery skew in the context of government advertising, come by and check it out!
I am excited to be at @facct.bsky.social this year presenting a new 📝 "Scrutinizing Index-Based Risk Assessments: A Case Study in NYC Decision-making for Heat Emergency Management" (work with Luke Boyce, @angelinawang.bsky.social , and @allisonkoe.bsky.social ). 🔗: dl.acm.org/doi/10.1145/... (1/10)
outrageously corny how happy and grateful I feel after PLSC, every single time. man do I love my job and the lovely geniuses it gives me proximity to. privacy odd ducks conference forever!!!!
Another one of @ahalterman.bsky.social and @katakeith.bsky.social's papers that I think should be cited more by CSS researchers: What is a protest anyway? Codebook conceptualization is still a first-order concern in LLM-era classification arxiv.org/abs/2510.03541
What is a protest anyway? Codebook conceptualization is still a first-order concern in LLM-era classification
Generative large language models (LLMs) are now used extensively for text classification in computational social science (CSS). In this work, focus on the steps before and after LLM prompting -- conce...
arxiv.org
I feel like I should be seeing this paper cited more often by CSS folks!!!
Wild Anna’s Archive bounty, reads like a heist. They want someone to front tens of thousands of dollars to buy Library of Congress files for a 3k bounty. I imagine a leak investigation would have a very short suspect list
Btw, did a bit of a rebranding of the substack. Will endeavor to post more there. h/t @dbateyko.bsky.social on the Trials & Errors name. Super fitting name for a group whose focus is both in reinforcement learning and in law/governance research. www.trialserrors.ai
Trials & Errors | Peter Henderson | Substack
Various news, thoughts, and findings on the intersection of law, policy, and artificial intelligence. Click to read Trials & Errors, by Peter Henderson, a Substack publication with hundreds of subscri...
trialserrors.ai
This is a challenging legal problem for NeurIPS (and other conference participants)! You might be wondering how this is possible given the First Amendment? I wrote a quick explainer on the current status quo of relevant First Amendment cases & law to get you up to speed. 🔗👇
I've started a "History of NLP" repo to store all of these resources. I don't have time to add everything yet, but I'll keep chipping away, and help is welcome. github.com/maria-antoni...
GitHub - maria-antoniak/history-of-nlp: a public, crowd-sourced bibliography about the history of natural language processing (nlp)
a public, crowd-sourced bibliography about the history of natural language processing (nlp) - maria-antoniak/history-of-nlp
github.com
I'm lecturing about the "History of NLP" this week. What should I include? Any favorite anecdotes, images, people, methods? Slides, books, papers, or talks for inspiration or grounding? I've been maintaining a small collection here: www.are.na/maria-antoni...
📣 Call for Contributions: LEXam-v2 – A Benchmark for Legal Reasoning in AI How well do today’s AI systems really reason about law? We’re building a global benchmark based on real law school & bar exams. 🧵 Full details, scope, and how to contribute in the thread 👇
One joy of growing as a scholar & doer has been the pleasure of being supported to pay attention to other people’s excellent work and amplify it. Next week I am publishing an article summarizing over 170 articles on AI + science + policy and tomorrow I get to email their authors to say thanks <3
It's remarkable how early Ford Foundation was to law and technology in the midcentury
this is what you see moments before going down a cyberspace and law rabbit hole
it finally happened (my 3090 overheated and emergency shut off)
Here’s the article. I’ve had more positive feedback on it than things I spent a year on. Apparently describing a problem that thousands of Trust and Safety people are seeing but also see the world ignoring is a good way to win hearts and minds :) www.lawfaremedia.org/article/the-...
The Rise of the Compliant Speech Platform
Content moderation is becoming a “compliance function,” with trust and safety operations run like factories and audited like investment banks.
lawfaremedia.org
After having such a great time at #CHI2025 and #FAccT2025, I wanted to share some of my favorite recent papers here! I'll aim to post new ones throughout the summer and will tag all the authors I can find on Bsky. Please feel welcome to chime in with thoughts / paper recs / etc.!! 🧵⬇️:
i am launching a magazine with @kevinbaker.bsky.social and the rest of the reboot collective on thursday at gray area! you should be there! open.substack.com/pub/reboothq...
I've arrived in the 🌁Bay Area🌁, where I'll be spending the summer as a research fellow at Stanford's RegLab! If you're also here, LMK and let's get a meal / go on a hike / etc!!
Well, this was a nice surprise. Download it while it’s hot! lsolum.typepad.com/legaltheory/...
Grimmelmann, Sobel, & Stein on Generative AI and Legal Interpretation
James Grimmelmann (Cornell Law School; Cornell Tech), Benjamin Sobel (Cornell University - Cornell Tech NYC), & David Stein (Vanderbilt University - Vanderbilt Law School) have posted Generative Misin...
lsolum.typepad.com
I am presenting a new 📝 “Bias Delayed is Bias Denied? Assessing the Effect of Reporting Delays on Disparity Assessments” at @facct.bsky.social on Thursday, with @aparnabee.bsky.social, Derek Ouyang, @allisonkoe.bsky.social, @marzyehghassemi.bsky.social, and Dan Ho. 🔗: arxiv.org/abs/2506.13735 (1/n)
I am so excited to be in 🇬🇷Athens🇬🇷 to present "A Framework for Auditing Chatbots for Dialect-Based Quality-of-Service Harms" by me, @kizilcec.bsky.social, and @allisonkoe.bsky.social, at #FAccT2025!! 🔗: arxiv.org/pdf/2506.04419