Lovely work led by Dharsh Kumaran and his team at Google DeepMind shedding new light on the metacognitive profile of LLMs We show that LLM confidence estimates are both stubborn and brittle all at once Paper: www.nature.com/articles/s42...
Competing Biases underlie Overconfidence and Underconfidence in LLMs - Nature Machine Intelligence
Kumaran et al. show that large language model (LLM) confidence is shaped by two competing biases: a choice-supportive bias that inflates confidence in initial answers, and a systematic overweighting o...
nature.com