FactAlign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment
Recasts black-box hallucination detection as knowledge-graph alignment, which also separates intrinsic from extrinsic hallucinations: 0.889 F1 on detection (WikiBio GPT-3) and 0.825 F1 on type classification (XSum), with no fine-tuning and no repeated sampling.