Reposted by Esteban Hernandez-Rivera
Lovely work led by Dharsh Kumaran and his team at Google DeepMind shedding new light on the metacognitive profile of LLMs
We show that LLM confidence estimates are both stubborn and brittle all at once
Paper: www.nature.com/articles/s42...
nature.com
Competing Biases underlie Overconfidence and Underconfidence in LLMs - Nature Machine Intelligence
Kumaran et al. show that large language model (LLM) confidence is shaped by two competing biases: a choice-supportive bias that inflates confidence in initial answers, and a systematic overweighting o...