Thomas Reed, and George Mason. “Hallucination Detection and Confidence Calibration for Large Language Model Outputs: Reproducible Experiments on HaluEval”. Journal of Artificial Intelligence Review 6, no. 4 (October 5, 2025): 1–17. Accessed October 10, 2026. https://learnedvertex.com/index.php/JAIR/article/view/321.