Skip to content

๐Ÿ‘ป Hallucination Detection

๐Ÿง  NeurIPS2026 ยท 1 paper notes

๐Ÿ“Œ Same area in other venues: ๐ŸŽž๏ธ ECCV2026 (20) ยท ๐Ÿ“ท CVPR2026 (33) ยท ๐Ÿ”ฌ ICLR2026 (40) ยท ๐Ÿ’ฌ ACL2026 (28) ยท ๐Ÿงช ICML2026 (21) ยท ๐Ÿค– AAAI2026 (15)

The Commit-Abstain Circuit: Why Language Models Hallucinate Instead of Abstaining

The paper localises a sparse Commit-Abstain Circuit (CAC) using the pre-generation commitโ€“abstain logit margin, identifies early commitment accumulation with insufficient late abstention correction, and trains a lightweight policy on component contributions that raises reported mean decision accuracy from 0.692 to 0.814, rather than demonstrating improved answer-content accuracy.