Most Multilingual benchmarks measure what models know, not what they can reliably do. Our #ACL2026 paper introduces functional benchmarks in six languages from English to Yoruba to test whether models can actually execute tasks across languages, not just answer fixed questions about them. 🧵1/n
Michelle L. Ding
@michelleding.bsky.social
researcher/organizer critically investigating how AI systems impact communities. cs phd @ brown cntr. she/her. 🌷 https://michelle-ding.github.io/ 💭 https://michellelding.substack.com/
Part IV: So what do we do? Don't overindex on one episode — that breeds the whack-a-mole regulation that's been harmful in other settings: @michelleding.bsky.social @harinisuresh.bsky.social arxiv.org/abs/2602.04759
How to Stop Playing Whack-a-Mole: Mapping the Ecosystem of Technologies Facilitating AI-Generated Non-Consensual Intimate Images
The last decade has witnessed a rapid advancement of generative AI technology that significantly scaled the accessibility of AI-generated non-consensual intimate images (AIG-NCII), a form of image-bas...
arxiv.org
Anthropic just shut access to Fable & Mythos, their most powerful models. I wrote up what happened + the policy and psychodrama behind it (with all the links & receipts) here — the thread below is the TL;DR 🧵 blog.geomblog.org/2026/06/the-...
The "fable" of Anthropic and the USG
News moves fast. 12 hours ago I was enjoying the demolition that the US put on Paraguay when I heard that Anthropic had shut down access to ...
blog.geomblog.org
On the one-year anniversary of EMPIRE OF AI, I am so, so excited to announce The AI Resist List, a new project that documents examples of resistance to the AI empires around the world 😍 airesistlist.org
Tech Policy Press fellow Petra Molnar highlights the AI Resist List: a global database documenting acts of resistance to the AI industry. From legal challenges and worker organizing to artistic interventions, the project seeks to challenge the “scale at all costs” development of AI.
The World Is Already Resisting AI. Now, There is a List to Prove It.
Petra Molnar spotlights the launch of the AI Resist List, documenting global challenges to AI expansion.
buff.ly
Fantastic thread by @michelleding.bsky.social about her new (and difficult) work on addressing deepfake abuse.
How do we stop playing whack-a-mole when it comes to deepfake abuse? 🧵⚠️
How do we stop playing whack-a-mole when it comes to deepfake abuse? 🧵⚠️
It's been a journey of nearly 3 years, but I'm very excited to announce the CNTR AISLE Portal! 🚀 cntr-aisle.org It’s a new way to review and evaluate the 1,000+ AI bills introduced in the U.S. over the last three years. Check out the Bill Library and our Profiles#AIPolicy #OpenData
CNTR AISLE
CNTR AISLE Portal
cntr-aisle.org
The Commodification of AI Sovereignty: Lessons from the Fight for Sovereign Oil (with Kate E. Creasey, Taylor Lynn Curtis, and @geomblog.bsky.social), is out now on arXiv: www.arxiv.org/abs/2601.11763.
We released a new report in partnership with the Center for Tech Responsibility at Brown University on how policymakers and researchers can better analyze AI legislation to protect our civil rights and liberties.
Making Sense of AI Policy Using Computational Tools | TechPolicy.Press
A new report examines how to use computational tools to evaluate policy, with AI policy as a case study.
techpolicy.press
Today on @indicator.media's free weekly briefing: The staggering impunity of xAI, which turned its abusive image generator on its own users in full view and has barely done anything to contain it.
Briefing: Grok brings nonconsensual image abuse to the masses
Plus: a new feature in the Meta Ad Library and a new Telegram investigation tool.
indicator.media
New post by @michelleding.bsky.social on resources for the Brown community in the aftermath of the shooting. open.substack.com/pub/michelle...
Caring for yourself and each other
Resources for the Brown community, friends, family, loved ones and how to support us
open.substack.com
@mantzarlis.com and folks at @indicator.media have done incredible reporting & investigation on the AI nudification ecosystem that I'm constantly citing in my research on AIG-NCII - appreciate all the work you do!
Today on Indicator: 2025 has been a banner year for AI nudifiers. I found another 9,000 ads on Meta since my last report, bringing the total for this year to 25,000. The top 10 nudifying websites got 10 million views in October.
Very glad to be a part of a new paper detailing how developers and developer platforms can prevent AIG-NCII, a form of image based sexual abuse that disproportionately harms women and girls. Thanks to all the collaborators and Max Kamachee & @scasper.bsky.social for leading this important project!
Did you know that one base model is responsible for 94% of model-tagged NSFW AI videos on CivitAI? This new paper studies how a small number of models power the non-consensual AI video deepfake ecosystem and why their developers could have predicted and mitigated this.
Thanks to collaborators! This was a really interesting paper for me to work on, and it took a special group of interdisciplinary people to get it done. Max Kamachee @r-jy.bsky.social @michelleding.bsky.social @ankareuel.bsky.social @stellaathena.bsky.social @dhadfieldmenell.bsky.social
ACM members/computing researchers who should be members interested in contributing should join the subcommittee's mailing list! One of our goals here is to build policy coalitions across institutions so we can do more as a collective 💪 and balance special interest groups.
@reniebird.bsky.social and I have just been appointed to co-Chair @TheOfficialACM's US Technology Policy Committee’s Subcommittee on AI and Algorithms. cs.brown.edu/news/2025/11...
@reniebird.bsky.social and I have just been appointed to co-Chair @TheOfficialACM's US Technology Policy Committee’s Subcommittee on AI and Algorithms. cs.brown.edu/news/2025/11...
Serena Booth And Suresh Venkatasubramanian Co-Chair ACM’s US Technology Policy Committee’s Subcommittee On AI And Algorithms
Brown CS faculty members Serena Booth and Suresh Venkatasubramanian have just been appointed to co-chair the AI and Algorithms Subcommittee, whose recent work includes responses to government RFIs, te...
cs.brown.edu
PSA: tips to protect yourself from scams on Signal. Every major comms platform has to contend w phishing, impersonation, & scams. Sadly. Signal is major, and as we've grown we've heard about more of these attacks--scammy people pretending to be something or someone to trick and abuse others. 1/
Today on @indicator.media: A first-of-its-kind audit of AI labels on major social platforms.
Tech platforms promised to label AI content. They're not delivering.
An Indicator audit of hundreds of synthetic images and videos reveals that platforms frequently fail to label AI content
indicator.media
Technologies like synthetic data, evaluations, and red-teaming are often framed as enhancing AI privacy and safety. But what if their effects lie elsewhere? In a new paper with @realbrianjudge.bsky.social at #EAAMO25, we pull back the curtain on AI safety's toolkit. (1/n) arxiv.org/pdf/2509.22872
arxiv.org
I wrote a (personal) blog post about my hopes and dreams for AI policy, my devastation after the US Election, and my process of picking myself off the floor by rebuilding an optimistic vision for AI scientists in government through education: simons.berkeley.edu/news/rebuild...
Rebuilding an Optimistic Vision for AI Policy
Recall November 6, 2024 — the day after the U.S. election. I was driving back to my home in Washington, DC, from Ohio with colleagues. I was heartbroken not because of the rebuke to my political party...
simons.berkeley.edu
💡We kicked off the SoLaR workshop at #COLM2025 with a great opinion talk by @michelleding.bsky.social & Jo Gasior Kavishe (joint work with @victorojewale.bsky.social and @geomblog.bsky.social ) on "Testing LLMs in a sandbox isn't responsible. Focusing on community use and needs is."
Have you or a loved one been misgendered by an LLM? How can we evaluate LLMs for misgendering? Do different evaluation methods give consistent results? Check out our preprint led by the newly minted Dr. @arjunsubgraph.bsky.social, and with Preethi Seshadri, Dietrich Klakow, Kai-Wei Chang, Yizhou Sun
Agree to Disagree? A Meta-Evaluation of LLM Misgendering
Numerous methods have been proposed to measure LLM misgendering, including probability-based evaluations (e.g., automatically with templatic sentences) and generation-based evaluations (e.g., with aut...
arxiv.org
🚨 New preprint! 🚨 Excited to share my work: An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies 🤖🗳️ I’ll be presenting this at @colmweb.org in the NLP4Democracy workshop! 🔗 arxiv.org/abs/2509.12577
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
In an era of increasing societal fragmentation, political polarization, and erosion of public trust in institutions, representative deliberative assemblies are emerging as a promising democratic forum...
arxiv.org
Hi #COLM2025! 🇨🇦 I will be presenting a talk on the importance of community-driven LLM evaluations based on an opinion abstract I wrote with Jo Kavishe, @victorojewale.bsky.social and @geomblog.bsky.social tomorrow at 9:30am in 524b for solar-colm.github.io Hope to see you there!
Third Workshop on Socially Responsible Language Modelling Research (SoLaR) 2025
COLM 2025 in-person Workshop, October 10th at the Palais des Congrès in Montreal, Canada
solar-colm.github.io
Very excited to be part of this new AI Institute that is being led by Ellie Pavlick @brown.edu and to be able to work with so many experts, including @datasociety.bsky.social www.brown.edu/news/2025-07...
Brown University to lead national institute focused on intuitive, trustworthy AI assistants
A new institute, based at Brown and supported by a $20 million National Science Foundation grant, will convene researchers to guide development of a new generation of AI assistants for use in mental a...
brown.edu
I'll be presenting a position paper about consumer protection and AI in the US at ICML. I have a surprisingly optimistic take: our legal structures are stronger than I anticipated when I went to work on this issue in Congress. Is everything broken rn? Yes. Will it stay broken? That's on us.
With their 'Sovereignty as a Service' offerings, tech companies are encouraging the illusion of a race for sovereign control of AI while being the true powers behind the scenes, write Rui-Jie Yew, Kate Elizabeth Creasey, Suresh Venkatasubramanian.
'Sovereignty' Myth-Making in the AI Race | TechPolicy.Press
Tech companies stand to gain by encouraging the illusion of a race for 'sovereign' AI, write Rui-Jie Yew, Kate Elizabeth Creasey, Suresh Venkatasubramanian.
techpolicy.press
Very excited to see this piece out in @techpolicypress.bsky.social today. This was written together with @r-jy.bsky.social and Kate Elizabeth Creasey (a historian here at Brown), and calls out what we think is a scary and interesting rhetorical shift. www.techpolicy.press/sovereignty-...
'Sovereignty' Myth-Making in the AI Race | TechPolicy.Press
Tech companies stand to gain by encouraging the illusion of a race for 'sovereign' AI, write Rui-Jie Yew, Kate Elizabeth Creasey, Suresh Venkatasubramanian.
techpolicy.press
So the EU AI Act passed. Companies have to comply. AI regulation is here to stay. Right? Right? FAccT 2025 paper with @r-jy.bsky.social and Bill Marino (not on bsky) 📜 incoming! 1/n arxiv.org/abs/2506.01931
Red Teaming AI Policy: A Taxonomy of Avoision and the EU AI Act
The shape of AI regulation is beginning to emerge, most prominently through the EU AI Act (the "AIA"). By 2027, the AIA will be in full effect, and firms are starting to adjust their behavior in light...
arxiv.org