Hannah Zhao and Yifan Zhang (2024) “Helpful or Harmful? Benchmarking Large Language Models as Therapy Tools Across Empathy, Specificity, and Safety”, Journal of Advanced Computing & Intelligent Systems, 4(7), pp. 93–109. doi:10.69987/.