Mark Whiting

@mark.whiting.me

Pareto.ai & CSS.seas.uPenn.edu → whiting.me

My team at Pareto.ai has been trying to understand LLM's potential for metacognition. As it turns out, they don't tend to have it beyond what they appear to have memorized awareness of from pre-training. Check out the details here: pareto.ai/blog/llm-met...

LLM Metacognition: Shared and Shallow?

Across 19 frontier models, metacognitive confidence on question and answer tasks tracks a shared difficulty heuristic with only a weak relationship to actual performance.

pareto.ai

Excited to see our work coming out + @joshnguyen.bsky.social & @duncanjwatts.bsky.social After establishing a means to study common sense in humans (and finding it rather limited — common sense is not so common) in a prior paper, we wondered if the same challenge faced language models. It does!

Josh Nguyen@joshnguyen.bsky.social · 8mo ago

Benchmarks of LLM common sense overwhelmingly rely on correct labels to report an accuracy score. But what if your "ground truth" genuinely differs from mine? In a new @pnasnexus.org paper, @duncanjwatts.bsky.social, @whiting.me and I explore the implications of this intriguing question. 🧵⤵️

What if technology didn’t feel so… hollow? Some friends and I just released a manifesto about a world where tech leaves us feeling nourished (along with an evolving list of theses about how we can build it) resonantcomputing.org

The Resonant Computing Manifesto

Technology should bring out the best in humanity, not the worst—a manifesto for resonant computing built on five principles that reject hyper-scale extraction for human flourishing.

resonantcomputing.org

Looking forward to reading this new PNAS paper "A framework for quantifying individual and collective common sense" doi.org/10.1073/pnas... Reminds of something I've always wondered...if there is a quantitative way to think about how Sherlock Holmes patches together his observations