Kate Sanders
@kesnet50.bsky.social
Researcher at Microsoft Copilot Tuning. Cal alum, Ph.D. @ JHU CLSP. #NLProc https://katesanders9.github.io/
This year's shared task allows you to submit for the retrieval track, generation track, or full RAG track on a challenging new collection of unedited ("raw") videos. Research Papers (Apr. 1) Shared Task (Apr. 20)
📹 + 🧠 + 📝 = 🔥 First call for MAGMaR 2026, the 2nd workshop on multimodal augmented generation via multimodal retrieval! If #RAG isn't hard enough for you, try multilingually and multimodally. Collocated with @aclmeeting in San Diego in July. nlp.jhu.edu/magmar/
MAGMaR Workshop
MAGMaR
nlp.jhu.edu
I will be at AAAI 2026 in Singapore next week! ✈️ I'm looking forward to seeing everyone's cool projects and discussing reasoning, post-training, and multimodality. Please reach out if you will be there and would like to connect.
Bye bye 2025, a divisive year, with many divisors: 3, 5, 9, 15, 25, 27, 45, 75, 81, 135, 225, 405, 675. Happy 2026 = 2*1013 Just two primes Cheers!
🚀 SynthTextEval, our open-source toolkit for generating and evaluating synthetic text data for high-stakes domains, will be featured at EMNLP 2025 as a system demonstration! GitHub: github.com/kr-ramesh/sy... Paper 📝: aclanthology.org/2025.emnlp-d... #EMNLP2025 #EMNLP #SyntheticData
GitHub - kr-ramesh/synthtexteval: SynthTextEval: A Toolkit for Generating and Evaluating Synthetic Data Across Domains (EMNLP 2025 System Demonstration)
SynthTextEval: A Toolkit for Generating and Evaluating Synthetic Data Across Domains (EMNLP 2025 System Demonstration) - kr-ramesh/synthtexteval
github.com
In honor of some new people coming from AI twitter, I finally updated my post to recommend For You over Discover.
I wrote something up for AI people who want to get into bluesky and either couldn't assemble an exciting feed or gave up doomscrolling when their Following feed switched to talking politics 24/7.
A company that believed it was in the verge of AGI or ASI wouldn’t capitulate to the government because it wouldn’t care about government contracts. They would soon BE the economy and the government would soon be capitulating to them.
The administration “is currently pressuring OpenAI and other AI companies to make their models more conservative-friendly.”
Keynote spotlight #4: the second day of COLM will close with @ghadfield.bsky.social from JHU talking about human society alignment, and lessons for AI alignment
Congratulations to Alane Suhr '22, a #CornellTech Ph.D. #alumni advised by associate professor Yoav Artzi, for receiving the prestigious 2022 @aaai.org / @acmsigai.bsky.social Doctoral Dissertation Award! Read more about the award here: aaai.org/about-aaai/a... @yoavartzi.com
AAAI/ACM SIGAI Doctoral Dissertation Award - AAAI
The AAAI/ACM SIGAI Doctoral Dissertation Award recognizes and encourages superior research and writing by doctoral candidates in AI.
aaai.org
Time for the world to install a gigawatt of solar power capacity 2004: A year 2010: ~ a month 2015: ~ a week Now: A day ourworldindata.org/data-insight... 🧪
🚨 Urban Stats 28.0.0 🚨 The mapper is now completely redesigned by me and @spudwaffle.bsky.social, allowing for much prettier looking maps and way more customization alongside significantly more options for geographies! See below for some of the examples of the maps you can create!
When reading AI reasoning text (aka CoT), we (humans) form a narrative about the underlying computation process, which we take as a transparent explanation of model behavior. But what if our narratives are wrong? We measure that and find it usually is. Now on arXiv: arxiv.org/abs/2508.16599
Humans Perceive Wrong Narratives from AI Reasoning Texts
A new generation of AI models generates step-by-step reasoning text before producing an answer. This text appears to offer a human-readable window into their computation process, and is increasingly r...
arxiv.org
So, what's the future of AI safety benchmarks? Jack's solution is "renewable benchmarks" that allows us to refresh and expand benchmarks with a single click!! x.com/jackjingyuz...
In our forthcoming paper, John Hummel and I ask what it would mean for a neural computing architecture such as a brain to implement a symbol system, and the related question of what makes it difficult for them to do so, with an eye toward the differences between humans, animals, and ANNs.
From Basic Affordances to Symbolic Thought: A Computational Phylogenesis of Biological Intelligence
What is it about human brains that allows us to reason symbolically whereas most other animals cannot? There is evidence that dynamic binding, the ability to combine neurons into groups on the fly, is...
arxiv.org
This paper is making the rounds: arxiv.org/abs/2506.21734 A tiny (27M) brain-inspired model trained just on 1000 samples outperforming o3-mini-high on reasoning tasks. #MLSky 🧠🤖
Interested in large-scale GPU optimization? Interested in how modern neural networks are being deployed to solve classical optimization problems? Writing a paper on these topics? Submit to the ScaleOPT workshop at NeurIPS! www.cvxgrp.org/scaleopt/#su...
ScaleOPT
cvxgrp.org
I'm recruiting MLEs @ #ACL2025! Reach out if you know folks interested in legal NLP, structured prediction, and full-time at a startup environment in NYC I'll also always chat about: • population-level inference on corpora • broad-coverage semantics • which café has the best Sachertorte in Vienna
My students and I are presenting three papers on Monday at #ACL2025 and this thread will recap them (including their videos).
“Wikipedia is this economic anomaly. In many ways, it’s sort of magical that people will just volunteer without explicit economic incentives to create artifacts that are meant to share knowledge with everyone in the world”
Taking off for Vienna #ACL2025! 🇦🇹 Excited to talk with people about transparent reasoning, multimodality, and fact verification. Stop by our multimodal RAG workshop on Friday 🔥🔥🔥 Please reach out if you want to grab coffee!
New Workshop on Multimodal Augmented Generation via MultimodAl Retrieval (MAGMaR) to be held at @aclmeeting.bsky.social ACL in Vienna this summer. We have a new shared task that stumps most LLMs - including ones pretrained on our test collection. nlp.jhu.edu/magmar/
The #ACL2025 #ACL2025NLP feed is up and running! It matches both hashtags and any posts from or mentions of @aclmeeting.bsky.social Pin it to your home 📌 and enjoy! bsky.app/profile/did:...
Juxtastat DAU update! Crazy how we've been >1000 every day for over a year now! Thank you all for all your support, and make sure to keep spreading the word!
🥳 🎉 ❤️ The ACL 2025 Proceedings are live on the ACL Anthology 🥰 ! We’re thrilled to pre-celebrate the incredible research 📚 ✨ that will be presented starting Monday next week in Vienna 🇦🇹 ! Start exploring 👉 aclanthology.org/events/acl-2... #NLProc #ACL2025NLP #ACLAnthology
Annual Meeting of the Association for Computational Linguistics (2025) - ACL Anthology
pdf bibProceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)Wanxiang Che | Joyce Nabende | Ekaterina Shutova | Mohammad Taher Pilehvar
aclanthology.org
This New Yorker piece is the most hopeful I've felt about the world in a long time. I had no idea solar was booming like this. And if you live in the same world as me, dominated by oil & gas guys maintaining that solar and wind are inefficient gimmicks, you might not've known some of this either.
anti-doomer sentence of the day: "Globally, roughly a third more power is being generated from the sun this spring than last" www.newyorker.com/news/annals-...
🔈When LLMs solve tasks with a mid-to-low resource input or target language, their output quality is poor. We know that. But can we put our finger on what breaks inside the LLM? We introduce the 💥 translation barrier hypothesis 💥 for failed multilingual generation with LLMs. arxiv.org/abs/2506.22724