Excited to share our paper Representational Difference Explanations (RDX) was accepted to #NeurIPS2025! 🎉RDX is a new method for model diffing designed to isolate 🔍 representational differences. 1/7
Stella Frank
@scfrank.bsky.social
Thinking about multimodal representations | Postdoc at UCPH/Pioneer Centre for AI (DK).
1/8 🧵 GPT-5's storytelling problems reveal a deeper AI safety issue. I've been testing its creative writing capabilities, and the results are concerning - not just for literature, but for AI development more broadly. 🚨
My Lab at the University of Edinburgh🇬🇧 has funded PhD positions for this cycle! We study the computational principles of how people learn, reason, and communicate. It's a new lab, and you will be playing a big role in shaping its culture and foundations. Spread the words!
🚀 DinoV3 just became the new go-to backbone for geoloc! It outperforms CLIP-like models (SigLip2, finetuned StreetCLIP)… and that’s shocking 🤯 Why? CLIP models have an innate advantage — they literally learn place names + images. DinoV3 doesn’t.
I wrote a short rant about what irks me when people anthropomorphize LLMs: addxorrol.blogspot.com/2025/07/a-no...
A non-anthropomorphized view of LLMs
In many discussions where questions of "alignment" or "AI safety" crop up, I am baffled by seriously intelligent people imbuing almost magic...
addxorrol.blogspot.com
📢I am hiring a Postdoc to work on post-training methods for low-resource languages. Apply by August 15 employment.ku.dk/faculty/?sho.... Let's talk at #ACL2025NLP in Vienna if you want to know more about the position and life in Denmark.
Postdoc in Natural Language Processing
employment.ku.dk
New paper hot off the press www.nature.com/articles/s41... We analysed over 40,000 computer vision papers from CVPR (the longest standing CV conf) & associated patents tracing pathways from research to application. We found that 90% of papers & 86% of downstream patents power surveillance 1/
Computer-vision research powers surveillance technology - Nature
An analysis of research papers and citing patents indicates the extensive ties between computer-vision research and surveillance.
nature.com
"Researching and reflecting on the harms of AI is not itself harm reduction. It may even contribute to rationalizing, normalizing, and enabling harm. Critical reflection without appropriate action is thus quintessentially critical washing."
Critical AI Literacy: Beyond hegemonic perspectives on sustainability open.substack.com/pub/rcsc/p/c...
Fallacy of the Day: Calling two different things by the same name doesn't make them the same (jingle) and calling the same thing by different names doesn't make them different (jangle) en.wikipedia.org/wiki/Jingle-... (this is going to be so useful for reviewing)
Jingle-jangle fallacies - Wikipedia
en.wikipedia.org
📯 Best Paper Award at CVPR workshop on Visual concepts for our (@doneata.bsky.social + @delliott.bsky.social) paper on probing vision/lang/ vision+lang models for semantic norms! TLDR: SSL vision models (swinV2, dinoV2) are surprisingly similar to LLM & VLMs even w/o lang 👀 arxiv.org/abs/2506.03994
I am excited to announce our latest work 🎉 "Cultural Evaluations of Vision-Language Models Have a Lot to Learn from Cultural Theory". We review recent works on culture in VLMs and argue for deeper grounding in cultural theory to enable more inclusive evaluations. Paper 🔗: arxiv.org/pdf/2505.22793
Check out our new paper led by @srishtiy.bsky.social and @nolauren.bsky.social! This work brings together computer vision, cultural theory, semiotics, and visual studies to provide new tools and perspectives for the study of ~culture~ in VLMs.
I am excited to announce our latest work 🎉 "Cultural Evaluations of Vision-Language Models Have a Lot to Learn from Cultural Theory". We review recent works on culture in VLMs and argue for deeper grounding in cultural theory to enable more inclusive evaluations. Paper 🔗: arxiv.org/pdf/2505.22793
as an extra take-away, this implies that our eval tends to be overly precision focused. we should really think of what we lose in terms of recalls, as this directly relates to what we miss out for whom when we build these large-scale, general-purpose models. (4/4)
🚀 We are excited to introduce Kaleidoscope, the largest culturally-authentic exam benchmark. 📌 Most VLM benchmarks are English-centric or rely on translations—missing linguistic & cultural nuance. Kaleidoscope expands in-language multilingual 🌎 & multimodal 👀 VLMs evaluation
Join us and revolutionize Life Science Lab Automation! 🎓🤖💉 I am hiring a Postdoc in Robotics and Computer Vision for Life Science Laboratory Automation, in Copenhagen, Denmark. Is that you? 🙋♀️ efzu.fa.em2.oraclecloud.com/hcmUI/Candid...
Postdoc in Robotics and Computer Vision for Life Science Laboratory Automation - DTU Electro
As part of a joint research collaboration between DTU and Novo Nordisk, we are looking for a postdoc to join our multidisciplinary research program focusing on the interplay between AI-Protein design,...
efzu.fa.em2.oraclecloud.com
Today we are releasing Kaleidoscope 🎉 A comprehensive multimodal & multilingual benchmark for VLMs! It contains real questions from exams in different languages. 🌍 20,911 questions and 18 languages 📚 14 subjects (STEM → Humanities) 📸 55% multimodal questions
We are looking for two PhD students at our institute in Munich. Both postions are open-topic, so anything between cognitive science and machine learning is possible. More information: hcai-munich.com/PhDHCAI.pdf Feel free to share broadly!
hcai-munich.com
📢Excited to announce our upcoming workshop - Vision Language Models For All: Building Geo-Diverse and Culturally Aware Vision-Language Models (VLMs-4-All) @CVPR 2025! 🌐 sites.google.com/view/vlms4all
BirdCLEF25: Audio-based species identification focused on birds, amphibians, mammals, and insects in Colombia. 👉 www.kaggle.com/competitions... @cvprconference.bsky.social @kaggle.com #FGVC #CVPR #CVPR2025 #LifeCLEF [1/4]
Thanks to these insects, we can now study environmental microplastics retrospectively. 🔍 Even before Duprat began his now famous experiments with caddisfly larvae, insects in the wild were already experimenting with plastic... 🐛 14/x
What a weekend to find Heinrich von Kleist's Erzählungen next to the skip, in which the first story is literally about a man wreaking murderous havoc because of the imposition of arbitrary trade tariffs en.wikipedia.org/wiki/Michael...
The world's 500 richest people saw their combined wealth fall by a combined $208 billion, the most since Covid, as tariffs sent markets into a tailspin
Billionaires Lose Combined $208 Billion in One Day From Trump Tariffs
The world’s 500 richest people saw their combined wealth plunge by $208 billion Thursday as broad tariffs announced by President Donald Trump sent global markets into a tailspin.
bloomberg.com
I’m excited to share our newest paper which is the first to analyze all of our in our TINTIN Corpus: 1,030 comics from 144 countries. We asked: How much are comic layouts influenced by the writing systems of their authors? www.sciencedirect.com/science/arti...
Would you present your next NeurIPS paper in Europe instead of traveling to San Diego (US) if this was an option? Søren Hauberg (DTU) and I would love to hear the answer through this poll: (1/6)
NeurIPS participation in Europe
We seek to understand if there is interest in being able to attend NeurIPS in Europe, i.e. without travelling to San Diego, US. In the following, assume that it is possible to present accepted papers ...
docs.google.com
Beautiful and motivated use of generated imagery 💙 wattenberger.com/thoughts/our... by @wattenberger.com More of this kind of thing please!
Our interfaces have lost their senses
wattenberger.com
Does anyone have a nice guide to digital privacy aimed at non-citizen residents traveling & returning to the US? I've been trying to check in with scientists I know with upcoming travel and am realizing many folks could use a 101 on removing biometric unlocking, doing a fresh iOS install, etc.