ENTITY
FaithCoT-Bench
FaithCoT-Bench
PulseAugur coverage of FaithCoT-Bench — every cluster mentioning FaithCoT-Bench across labs, papers, and developer communities, ranked by signal.
Total · 30d
1
2 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D
1 day(s) with sentiment data
RECENT · PAGE 1/1 · 2 TOTAL
-
Research: CoT unfaithfulness detection fails on incorrect model answers
A new research paper published on arXiv explores the unfaithfulness of Chain-of-Thought (CoT) reasoning in large language models. The study, titled "Two Regimes of Chain-of-Thought Unfaithfulness: Behavioral Detection F…
-
New framework detects unfaithful chain-of-thought reasoning in LLMs
Researchers have developed a new framework called CIE-Scorer to detect when a large language model's chain-of-thought (CoT) reasoning does not accurately reflect its internal decision-making process. This method combine…