Researchers have developed I-CALM, a framework designed to improve the selective answering capabilities of large language models (LLMs). This method encourages LLMs to abstain from answering questions when they are likely to be incorrect, thereby reducing the rate of false answers while maintaining correct responses. I-CALM achieves this by eliciting confidence, defining answer/abstain payoffs, and guiding models toward truthfulness and humility, without requiring model retraining or access to internal states. AI
IMPACT Improves LLM reliability by enabling them to express uncertainty and abstain from answering when confidence is low.
RANK_REASON The cluster contains an academic paper detailing a new framework for LLMs. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →