๐ป Hallucination Detection¶
๐ง NeurIPS2026 ยท 1 paper notes
๐ Same area in other venues: ๐๏ธ ECCV2026 (20) ยท ๐ท CVPR2026 (33) ยท ๐ฌ ICLR2026 (40) ยท ๐ฌ ACL2026 (28) ยท ๐งช ICML2026 (21) ยท ๐ค AAAI2026 (15)
- The Commit-Abstain Circuit: Why Language Models Hallucinate Instead of Abstaining
-
The paper localises a sparse Commit-Abstain Circuit (CAC) using the pre-generation commitโabstain logit margin, identifies early commitment accumulation with insufficient late abstention correction, and trains a lightweight policy on component contributions that raises reported mean decision accuracy from 0.692 to 0.814, rather than demonstrating improved answer-content accuracy.