🚨 Out now in TICS How can language models help cognitive science? @ruimata.bsky.social & I outline 5 uses: mapping research fields, formalizing theories, cleaning up constructs & measures, predicting behavior across tasks, and capturing environmental variation. 🔗 doi.org/10.1016/j.ti...
Julian Berger
@officialberger.bsky.social
Interested in the decision making of humans and machines. Postdoc SDU, previously Max Planck Institute for Human Development, soon UPF
Psychological Science is running a special issue on "Advancing Psychological Science using Large Language Models", co-edited by me and @ruben.the100.ci. Check out the calls for submissions below! www.psychologicalscience.org/publications...
Special Issue Call for Submissions: “Advancing Psychological Science using Large Language Models”
Editors: Jamie Cummins and Ruben C. ArslanPsychological Science invites submissions for a special issue on psychological science involving large language models. We welcome theoretically and empirical...
psychologicalscience.org
Happy to support reproducibility efforts whenever we can. Reach out if you want to use rigor.me at scale or in teaching!
Reminder: tomorrow's the Reproducibility Hackathon @edfuturesinstitute.bsky.social! 💻🎉 Big thanks to our pals @mpib-berlin.bsky.social for letting us beta-test their new tool, rigor.me. We'll be reproducing a paper and testing our own code's reproducibility. Open science, one step at a time 🔬
Reminder: tomorrow's the Reproducibility Hackathon @edfuturesinstitute.bsky.social! 💻🎉 Big thanks to our pals @mpib-berlin.bsky.social for letting us beta-test their new tool, rigor.me. We'll be reproducing a paper and testing our own code's reproducibility. Open science, one step at a time 🔬
Multiverse analysis, abdication of responsibility and manufacturing of doubt: I have written on some downsides I see with multiverse analysis (which I like in principle): arxiv.org/abs/2607.14623 I was inspired/provoked to write it by @dingdingpeng.the100.ci (thanks!)
Today we launch the first stable release of RegCheck: v1.0.0. RegCheck makes it easier and quicker to compare study registrations to published papers for consistency - something we know is important in principle, but rarely done in practice. A 🧵 on what's new: regcheck.app
RegCheck
RegCheck is an AI tool to compare preregistrations with papers instantly.
regcheck.app
Research on cognitive offloading × AI is booming and getting lots of public attention. Unfortunately, some (or even many) of it raises serious credibility concerns. Like this paper, which others and I recently commented on via @pubpeer.com: pubpeer.com/publications...
PubPeer - AI Tools in Society: Impacts on Cognitive Offloading and the...
There are comments on PubPeer for publication: AI Tools in Society: Impacts on Cognitive Offloading and the Future of Critical Thinking (2025)
pubpeer.com
Every once in a while you find a paper that does a lot with a simple approach. Today's is from @natematias.bsky.social, Cassidy Waldrip, and @davidlazer.bsky.social. They simply ask how much computational social science can be reproduced? #metascience. journalqd.org/article/view...
View of Commercial Determinants of Replication Infeasibility in Computational Social Science
journalqd.org
Come work with me, Janna, Jonas, and Bart at Leiden University :). #MentalHealth #Psychology #PsychSciSky #academicsky
1/ We have a PhD position open on 'Mental Disorders as Harmful Stable States'. If you know students interested in the intersection of mental health & statistical modeling (EMA & time series), please encourage them to apply. www.academictransfer.com/nl/jobs/3621...
We have some slots available tomorrow and over the weekend if folks want to test how reproducible their work is! Get a free reproducibility check of your paper: rigor.me I would especially welcome some neuro & ML & econ folks to test this. Throw any weird data at it. Much appreciated ❤️
Rigor
rigor.me
Sharing our latest endeavour here. How reproducible is your paper? @philipjakobbln.bsky.social and I built rigor.me to ease the burden of computational reproducibility. If you provide a paper, data and code, we execute it and tell you what works (and what fails). Beta is available now: rigor.me
🚨 Now out in PNAS In decision research, "talk is cheap." We show it isn't. Using LLMs to analyze participants' free-text explanations of their choices, we find verbal reports are a rich, scalable window into how people actually decide. 🔗 www.pnas.org/doi/10.1073/... Led by Kamil Fulawka
Damn! I ran my just-published paper through it and the review was awesome, if a bit humbling. I agree with most of it, and I would have addressed several of the critiques if I had this before submitting. Great stuff!
Usability 10/10 (though of course more speed = more better but I understand this takes time and compute isn't infinite). I will have to carefully check the discrepancies but I immediately understood some of them. Fortunately there was nothing major this time :)
I tried this on one of our papers from 2021. It did a fantastic job! For example (see screenshot) it verified our reported sample size & response rate, and about 80 other reported numbers/statistics with descriptions of where, why, and how some numbers differed.
Sharing our latest endeavour here. How reproducible is your paper? @philipjakobbln.bsky.social and I built rigor.me to ease the burden of computational reproducibility. If you provide a paper, data and code, we execute it and tell you what works (and what fails). Beta is available now: rigor.me
Like the wallet inspector, but for your data and code. (Kidding, this is very cool and useful)
Sharing our latest endeavour here. How reproducible is your paper? @philipjakobbln.bsky.social and I built rigor.me to ease the burden of computational reproducibility. If you provide a paper, data and code, we execute it and tell you what works (and what fails). Beta is available now: rigor.me
Sharing our latest endeavour here. How reproducible is your paper? @philipjakobbln.bsky.social and I built rigor.me to ease the burden of computational reproducibility. If you provide a paper, data and code, we execute it and tell you what works (and what fails). Beta is available now: rigor.me
Rigor
rigor.me
AUTOMATING COMPUTATIONAL REPRODUCIBILITY My colleague @philipjakobbln.bsky.social and I are currently engaged in a research project where we reproduce scientific results en masse. To that end, we built rigor.me, a platform for automatically reproducing papers using agents 🧵
AUTOMATING COMPUTATIONAL REPRODUCIBILITY My colleague @philipjakobbln.bsky.social and I are currently engaged in a research project where we reproduce scientific results en masse. To that end, we built rigor.me, a platform for automatically reproducing papers using agents 🧵
Our paper "Recognising and mitigating LLM Pollution in online behavioural research" is now officially published in Nature Communications. Congratulations to our team members Raluca Rilla, Tobias Werner, Hiromu Yakura, Anne-Marie Nussberger www.nature.com/articles/s41...
Recognising and mitigating LLM Pollution in online behavioural research - Nature Communications
Online behavioural research faces a growing methodological and epistemic threat as participants increasingly rely on large language models: LLM Pollution. Amid accumulating empirical evidence of conta...
nature.com
New preprint led by @lucasmolleman.bsky.social investigates social information use across the human lifespan. Testing >40,000 participants aged 6-80 in museums in Berlin and Japan. psyarxiv.com/hec96
🧵 Thread: Does the internet threaten democracy? I argue that it does in my latest piece in Science: doi.org/10.1126/scie... Here's the short version. 1/9 1/10
The architecture of the internet creates risks for democracy
Will democracy survive the internet? Do we need to choose between Facebook’s surveillance capitalism or democracy? Layered lines of evidence can inform questions like these. When considered together, the evidence gives rise to a concerning picture, as summarized in a recent report for the European Commission that I co-led.
doi.org
Postdoc position @unimarburg.bsky.social in the project: "Bridging the Gap Between Verbal Psychological Theories & Formal Statistical Modeling with Large Language Models" (funded by @volkswagenstiftung.de) 📅Start: 01.10.2026 |⏳3 years 🔗 Job posting: uni-marburg.de/78NrWT Thanks for sharing!
My team, Minsu Park and I just launched 12points.science. It’s a citizen science platform and app where you can rate, rank, and comment on Eurovision entries with your friends. There is a deeper research goal behind the glitter 🎤🧪 #Eurovision #OpenData
Can you boost your AI review scores by asking an LLM to rewrite your paper? Yes! We call it paper laundering Our @icmlconf.bsky.social spotlight paper argues current AI reviewers aren't ready to automate peer review, and outlines what a science of peer review automation should look like 🧵👇 #ICML2026
People in the AI world seem to adore the System 1/2 dichotomy. I love this quote from Gerd Gigerenzer about the concept:
Reminder: when authors observe Cohen's d = 4, no they didn't. Article now retracted: www.nature.com/articles/s41... Critique: osf.io/preprints/ps... Blog on plausibility of effect sizes: trustworthy.scientific.claims/posts/if-res...
RETRACTED ARTICLE: The effect of ChatGPT on students’ learning performance, learning perception, and higher-order thinking: insights from a meta-analysis - Humanities and Social Sciences Communication...
Humanities and Social Sciences Communications - RETRACTED ARTICLE: The effect of ChatGPT on students’ learning performance, learning perception, and higher-order thinking: insights from a...
nature.com
Big news! The 14th ACM Collective Intelligence Conference (CI 2026) will be held September 27-30, 2026, at the Virginia Tech Institute for Advanced Computing near Washington, DC. I'm excited to be a program co-chair this year. ci.acm.org/2026/ Papers and abstracts due 8 June. Check it out!
2026 ACM Collective Intelligence Conference
ci.acm.org
1/ "Silicon samples" are becoming more and more common in research and polling. One problem: depending on the analytic decisions made, you can basically get these samples to show any effect you want. The updated version of this preprint is now online! THREAD🧵 arxiv.org/abs/2509.13397
The threat of analytic flexibility in using large language models to simulate human data
Social scientists are now using large language models to create "silicon samples": synthetic datasets intended to stand in for human respondents. However, producing these samples requires many analyti...
arxiv.org
Can large language models stand in for human participants? Many social scientists seem to think so, and are already using "silicon samples" in research. One problem: depending on the analytic decisions made, you can basically get these samples to show any effect you want. THREAD 🧵
I’m hiring a PhD student! The candidate will work alongside @zefreeman.bsky.social, who is joining our research group as postdoc. jobs.unibe.ch/job-vacancie...
PhD Student in Meta-Science and Clinical Psychology - Universität Bern
Universität Bern is looking for PhD Student in Meta-Science and Clinical Psychology
jobs.unibe.ch
Do you want to do a PhD in Judgment and Decision Making in two beautiful European locations on human-AI interaction? Well have I got news for you. @bahniks.bsky.social and I are recruiting a candidate to start around September 2026. See here: decisionlab.vse.cz/english/we-a...
We are hiring
PhD Position in Judgment & Decision Making At a Glance Topic: Human-AI Interaction & Cognitive Biases Locations: VŠE Prague (2 years) + Maastricht University (2 years) Funding: Fully funded (4-year pr...
decisionlab.vse.cz