Andrew Gordon

@andrewgordon.bsky.social

Staff Researcher (Behavioural Science) @Prolific Previous dabbler in the cognitive neuroscience of misinformation Once bitten by the worlds rarest goose Was once bitten by the worlds rarest goose

We have a new preprint that underscores some key claims here: even if one *can* design an agent that gets through a survey fine, it doesn't follow that such agents are undetectable or common. We find that they are far from common! Preprint link in thread👇

Association for Psychological Science@psychscience.bsky.social · 4mo ago

Psychological scientists are determined to figure out the best practices for online surveys, who—or what—is behind bad data, and how to best protect surveys from the new and emerging threat of #AI. @justinsulik.bsky.social

Fantastic example of researchers working together (and the utility of rebuttals to published work). I think we all agree that this is an area we need to invest time in, but we also need to be very careful that conclusions/interpretations are warranted from the data we collect.

Christoph Strauch@cstrauch.bsky.social · 5mo ago

We recently warned of bots in online behavioral research. @achetverikov.bsky.social showed there is no evidence for that in our @joinprolific.bsky.social data - but that doesn't mean we're safe. Agentic AI can do behavioral tasks through prompting alone. Reply & videos: osf.io/3cztr/overview

AI will soon expose that most published research is mediocre, irrelevant (or simply bullsh*t). Most scientists know that, but the public doesn't... What's it going to do for trust in science when they find out - how do we plan for it? 🤔 Great article by @naomioreskes.bsky.social shorturl.at/i2Hn5

AI will soon be able to audit all published research – what will that mean for public trust in science?

An AI audit of scientific research would likely expose some fraud and widespread inconsequential work. But we need to be careful it doesn’t discredit science in general.

theconversation.com

Plot twist: What if AI's best future might not be the insanely profitable one 🤔 Great article arguing that instead of trillion-dollar empires, we could get free, open-source models that are just "good enough" for most people. Would that be so bad? theconversation.com/generative-a... #AI #tech

Generative AI might end up being worthless — and that could be a good thing

GenAI does some neat, helpful things, but it’s not yet the engine of a new economy — and it might not ever be.

theconversation.com

📊 Our new @joinprolific.bsky.social AI User Experience Leaderboard is live! AI systems ranked by real human experience, not just technical metrics. Using Census-based sampling + MRP for results that represent the general public. Check it out here: huggingface.co/spaces/nlpet...

UX Leaderboard - a Hugging Face Space by nlpetprolific

Leaderboard of LLMs based on detailed human feedback

huggingface.co

Prolific@joinprolific.bsky.social · last yr.

📊 Introducing Prolific’s AI user experience leaderboard—the most reliable framework for evaluating AI against human preferences. We’re thrilled to share our benchmark on @hf.co, which assesses how well language models handle real-world tasks based on user experiences 👉 www.prolific.com/leaderboard

How does the US public feel about the incoming #Trump administration? We asked a representative sample of 1,938 US adults on Prolific "Are you hopeful or fearful ahead of President-elect Trump coming back to office?" 🇺🇸 Overall 46% were fearful and 41% hopeful. But stark differences in subgroups 👇

Bild