Brian Heseung Kim

@brhkim.bsky.social

Data Scientist & Education Policy Researcher // Building rigorous, open-source AI trainings and tooling for researchers and non-profits to use AI responsibly and critically // PhD + MPP @UVA // www.brhkim.com // www.openaugments.org

I've run intensive workshops on best practices for responsible and rigorous AI adoption for well over 500 researchers+academics over the past few months. The Q&A after my workshops are some of the most fun and productive conversations I've had, so I figure: let's make this its own regular thing!

Bild

AI can help us think *more* critically and *more* deeply, not less, and it's urgent that we start cultivating that practice for ourselves and our students. Lessons from a former English teacher battling SparkNotes: daafguide.substack.com/p/on-rethink... #EduSky #AIinEducation #HigherEd #AcademicSky

On rethinking coursework expectations in the era of AI: Raising the floor and the ceiling

What we can learn from high school English teachers and frontier AI model benchmarking

daafguide.substack.com

I've been running workshops on AI for academics across the country the past few months, and the same question has come up in every single room: So... what are we going to do about teaching? No easy answers, but I want to offer a few distilled suggestions: daafguide.substack.com/p/on-rethink...

On rethinking coursework expectations in the era of AI: Raising the floor and the ceiling

What we can learn from high school English teachers and frontier AI model benchmarking

daafguide.substack.com

Added GLM 5.2 to DAAFBench, and it's so good it required a complete rewrite of the key takeaways and recs. It is utterly indistinguishable from Opus 4.5/4.6/4.8 at 25-33% of the cost and open-weight. What a wild week. Kimi K2.7 Coder also added in, but frankly unimpressive!

Bild
Brian Heseung Kim@brhkim.bsky.social · 2mo ago

I ran 17 different frontier models through 51 different tests (2500 total runs!) to examine: How well do different AI models handle the complexities of rigorous quantitative data analysis workflows? Excited to introduce DAAFBench: Orchestration! daaf.openaugments.org/bench

🥳 If you’ve been wondering how to use Claude Code as a quantitative researcher/social scientist of any kind: I’ve *finally* made a very nice, very accessible, and very informative homepage for the Data Analyst Augmentation Framework (DAAF), and I think you'll wanna take a look!

🥳 DAAF v2.1.0 is live today! With it, I feel strongly that DAAF is now finally the best, safest, AND easiest (+free!) way to get started w/ Claude Code for *any* data work. v2.1.0 was focused on three core features for quality of life and ease of use: daafguide.substack.com/p/daaf-v210-...

DAAF v2.1.0: The Frictionless Update

In my very biased opinion: DAAF is now the best AND easiest way to get started using Claude Code for researchers

daafguide.substack.com

🙌 New explainer article! TLDR: Everyone needs to be investing in better logging+monitoring to track model adherence for their workflows, because we can't assume any two models will follow instructions the same way given how weird the AI frontier is right now daafguide.substack.com/p/opus-47-la...

The Opus 4.7 launch fiasco as a crucial reality check for anyone building with AI in 2026

Do you really know what your AI agents are doing right now?

daafguide.substack.com

🥳 It's been a month since I launched DAAF, the Data Analyst Augmentation Framework... which means now is great time to celebrate the launch of DAAF v2.0.0 and re-introduce you all to a more useful, usable, and flexible tool for anyone analyzing data in their work!! www.youtube.com/watch?v=747r...

DAAF v2.0.0 -- Responsible, Rigorous, and Reproducible AI-empowered Data Analysis with Claude Code

YouTube video by Brian Heseung Kim (brhkim)

youtube.com

🥳 It's been a month since I launched DAAF, the Data Analyst Augmentation Framework... which means now is great time to celebrate the launch of DAAF v2.0.0 and re-introduce you all to a more useful, usable, and flexible tool for anyone analyzing data in their work!! www.youtube.com/watch?v=747r...

DAAF v2.0.0 -- Responsible, Rigorous, and Reproducible AI-empowered Data Analysis with Claude Code

YouTube video by Brian Heseung Kim (brhkim)

youtube.com

LLM AI assistants will always be at risk of hallucinating/sycophancy/lying. Can they still be useful for accelerating good research? Yes! But we need a *lot* of guardrails. If you've wanted to learn how to use tools like Claude Code to *responsibly* accelerate quantitative research...

Bild

As AI tools for research proliferate and research production accelerates, publishing is going to need to split into two tracks in the near future: that which is mechanically reproducible and verifiable, and basically everything else.

BildBild

Excited to be spending the rest of the week at AEFP 2026! Lots of exciting opportunities to connect if you want to know where to find me: 1. Running a professional development session on Friday morn to help folks target job applications to non-research organizations for "embedded" researcher roles

BildBildBild

Brazenly irresponsible projects like this are exactly what drove me to develop DAAF vocally centering the human expert. If we don't form strong, principled positions about what good AI-empowered research *can* and *should* look like, we wind up getting dragged around by people who clearly...

Bild

How can we get to a more optimistic AI-empowered future for academia and science? Step 1: As AI allows us to accelerate the production of more and more scientific claims, we *must* center transparency and reproducibility as core requirements for all AI research tooling #academicchatter #econsky

Bild
Brian Heseung Kim@brhkim.bsky.social · 5mo ago

AI-empowered research is here, and it's going to break the peer review system. What can we actually do about it? daafguide.substack.com/p/a-better-a... #academicsky #highered #academicchatter #PhDChat #phdsky #econsky

Disappeared Tufts Human Dev. PhD student Rumesya Ozturk is also a graduate of Teachers College - she’s a dev. psychologist studying children’s media & prosocial development. She also bakes without recipes and binge-watches cartoons. She is our colleague & she was abducted on the street w/ our tax $.

BildBild