Detecting Hallucination in LLMs: Tracing the Topological Signatures of Impaired Context Sharing
arxiv.org原文 ↗
论文把幻觉检测建立在注意力图拓扑上,用 Forman-Ricci 曲率定位信息瓶颈,并结合半局部与全局流特征做单次前向判断。跨多个模型和两个检测基准时,它超过已有注意力及多回答基线;最后一层的自注意过度、检索扩散和 over-squashing 是反复出现的结构信号,给可解释监测提供了比 token 概率更具体的切口。
–浏览
评论 · Comments