Thomas Reed, and George Mason. “Hallucination Detection and Confidence Calibration for Large Language Model Outputs: Reproducible Experiments on HaluEval”. Journal of Artificial Intelligence Review, vol. 6, no. 4, Oct. 2025, pp. 1-17, https://doi.org/10.69987/.