arXiv Artificial Intelligence

I-CALM: Incentivizing Confidence-Aware Abstention for LLM Selective Answering

I-CALM: Incentivizing Confidence-Aware Abstention for LLM Selective Answering

Quick summary

arXiv:2604.03904v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often produce confident but incorrect answers, in part because standard evaluation incentives reward guessing over expressing uncertainty. We study epistemic abstention for factual questions with verifiable answers, where the goal is to improve selective answering, making LLMs abstain when they are likely to be wrong while preserving correct answers. Inspired by human behavioral decisions in question answering, we introduce I-CALM, a prompt-level framework for black-box LLMs. I-CALM combines elicited verbal

Key takeaways

  • arXiv:2604.03904v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often produce confident but incorrect answers, in part because standard evaluation incentives reward guessing over expressing uncertainty.
  • We study epistemic abstention for factual questions with verifiable answers, where the goal is to improve selective answering, making LLMs abstain when they are likely to be wrong while preserving correct answers.
  • Inspired by human behavioral decisions in question answering, we introduce I-CALM, a prompt-level framework for black-box LLMs.

Why it matters

“I-CALM: Incentivizing Confidence-Aware Abstention for LLM Selective Answering” highlights the need for repeatable measurement rather than a single impressive demonstration. Independent validation across datasets and clearly stated limitations determine whether a result can guide product decisions.

Kaynak sitede devamını oku: arXiv Artificial Intelligence ↗