Mattan S. Ben-Shachar

@mattansb.msbstats.info

Statistics lecturer | Freelance statistical consultant & research analyst | #rstats dev @easystats.github.io home.msbstats.info (He/Him)

My 19-year-old asked me if I found my job hard. I thought it was an interesting question and I had to think about it. After discussing with a colleague, he created the best analogy. He said our kind of work is “hard like an escape room with friends”, not “hard like hard math.” I think thats right. 

my annoyance of bayesian modeling is mainly due to selection bias - I approach it when there's no other viable alternative, so i always encounter the painstaking experience of needing to iterate and refine with every fit taking upwards of 30 minutes

The ability to properly define a research question is more important today than ever before. Describe a vague research question and your data and 10 times out of 10 Claude WONT ask for any clarification before it suggests 19 different things you can *try*.

Sean Mackinnon@seanpmackinnon.bsky.social · 3y ago

Stats consulting is constantly like: Them: I want you to run (complex stats) Me: OK, what's your research question though? Them: ...A (complex stat)? Me: Research question? Them: You know, like the stats in this journal article. Something reviewers will like.

Goose chasing meme.

"What's your research question?"

Then angry goose chasing yelling the same phrase.

Made a detailed analysis plan for a client to run. A week later they email me results to review. They don't match our discussed plan at all. What happened? They didn't understand some of my suggestions - did they email me to get clarifications? No - they asked Claude to "fill in the blanks".

I am begging you to talk to other human beings directly, and not through a Claude layer. This goes doubly when I ask you, directly, to talk to me not through a Claude layer. Please.

A client thought up some BS with Claude and sent me the full report + analysis code, all 100% written and run and interpreted by Claude. I gave a detailed response explaining why it was trash, which my client (for laughs) forwarded to Claude. Claude was not happy being criticized.

Bild

Me to students: in this course we will not be using the ~ notation in @mc-stan.org for reasons that will be made clear later. Me, later: see when estimating the marginal likelihood we need all the normalizing constants, so be wary of LLMs that like using ~ notation. Students' assignmens: ~~~~~~~~

Scary stories about teens sell. The data tells a different story. After decades studying adolescent mental health, here's what I found: This isn't an anxious generation. It's a resilient one. Let's start treating them that way. go.ted.com/candiceodgers

What we're getting wrong about teens and tech

For years, the warning has been: smartphones are destroying a generation. But developmental psychologist Candice Odgers says that decades of data on teens tells a different story — violence, alcohol u...

go.ted.com

i see a particular grievance regularly on here: "a paper I reviewed came back to me at a diff journal. the authors didn't address my concerns at all." seriously what are you doing accepting to review the same paper? let people breathe & have another chance. let them decide what's worth addressing.

Matt Weiner@mattweiner19.bsky.social · 2w ago

will the next reviewer have the same objection? if you revise to address the objection, will the next reviewer say "I don't see why the paper spends so much time addressing this objection that no one could possibly have"? who knows

Peer review is so random bro. Got 2 papers under review. One is clearly an excellent piece, splashy and well-argued. It's on its 3rd round of revisions with a reviewer who doesn't understand it. Other paper is shorter, "worse" (more narrowly focused). Just got the easiest first round R&R. Dice rolls

Since paper here is nested in journal, those 2 sources of variance can be IMO combined to represent between paper variability. And even w/ this, we get a generalizability score (reliability) of 0.31 for 2 judges and 0.40 for 3 judges. You'd need 11 (!) judges to get a reliability > 0.7! 1/

Alejandro Montenegro@aemonten.bsky.social · 2w ago

"Here we show, in a large post-publication peer review database, that research assessment is driven more by differences between evaluators than by difference in the evaluated research. " arxiv.org/abs/2607.09783

New preprint w/ Leo Tiokhin & @lakens.bsky.social! Registered Reports (RRs) are great for science b/c publication is guaranteed before results are known, reducing publication bias & QRPs. This feature supposedly also makes them great for *scientists*. But is that really true? 1/ tinyurl.com/4sy993rr

Career incentives make Registered Reports unattractive under realistic academic conditions

Registered Reports are a publication format designed to reduce publication bias by guaranteeing publication before results are known. This guarantee is considered attractive in the prevailing ‘publish...

tinyurl.com

Descriptive studies overadjusting for a dozen variables. Seemingly causal studies missing major confounders or adjusting for mediators. We recently received an R&R where the reviewer asked why we didn't adjust for confounding by a set of mediating factors in our descriptive study. #episky 2/3