1.
Thomas Reed, George Mason. Hallucination Detection and Confidence Calibration for Large Language Model Outputs: Reproducible Experiments on HaluEval. JAIR [Internet]. 2025 Oct. 5 [cited 2026 Oct. 10];6(4):1-17. Available from: https://learnedvertex.com/index.php/JAIR/article/view/321