Functional vs. phenomenological empathy in large language models: Rethinking artificial empathy through experimental evidence

Authors

DOI:

https://doi.org/10.29038/eejpl.2026.13.1.mug

Keywords:

psycholinguistics, Large Language Models, functional empathy, phenomenological empathy

Abstract

Although large language models (LLMs) are increasingly used in emotionally sensitive contexts, it remains unclear whether their responses reflect genuine empathy or functional simulation. This study employed a mixed-methods experimental design to compare empathetic responses from human participants (n = 100) with those generated by LLMs under three conditions: text-only, anthropomorphic cue, and multimodal with cue. Using ten standardized emotional vignettes, 200 independent raters evaluated all responses with the Consultation and Relational Empathy (CARE) measure. Quantitative results showed that LLM-generated responses were consistently rated as more empathetic than human responses, with empathy scores increasing across enhanced design conditions. Qualitative analysis using the Empathic Communication Coding System (ECCS) revealed that LLMs relied heavily on surface validation and supportive strategies, while demonstrating limited contextual probing and a lack of experiential grounding compared to human responses. These findings highlight a critical distinction between perceived empathy and phenomenological empathy. While LLMs can effectively simulate empathic communication and achieve high ratings due to consistency and alignment with social expectations, they lack the contextual depth and experiential understanding that characterize human empathy. This study contributes to the literature by moving beyond performance-based evaluations and clarifying the structural differences between human and artificial empathy. The results have important implications for the use of LLMs in healthcare, education, and counseling, where perceived empathy may enhance user experience but also raises concerns about over-trust and anthropomorphic attribution.

CRediT Statement

Ahmad Mugableh: Conceptualization, Methodology, Software, Validation, Formal Analysis, Investigation, Writing – Original Draft, Visualization, Supervision, Project Administration; Hissah Mohammed Alruwaili: Resource, Data Curation, Writing – Reviewing & Editing.

Disclosure Statement

The authors reported no potential conflicts of interest.

Generative AI Statement

The authors used an AI-assisted language tool during the preparation and revision of the manuscript to support language refinement, clarity improvement, formatting consistency, and editorial revision. The AI tool was not used to generate original data, conduct statistical analyses, interpret findings independently, or make scholarly decisions regarding the study’s design, methodology, results, or conclusions. All conceptual development, data analysis, interpretation, and final manuscript decisions were performed and verified by the authors, who take full responsibility for the manuscript's content.

Downloads

Download data is not yet available.

Author Biography

  • Ahmad Mugableh *, Department of English, Jouf Univeristy, Saudi Arabia

    * Corresponding author, Email: [email protected]

References

Batson, C. D. (2009). These things called empathy: Eight related but distinct phenomena. In J. Decety & W. Ickes (Eds.), The social neuroscience of empathy (pp. 3–15). Boston Review. https://doi.org/10.7551/mitpress/9780262012973.003.0002

Braun, V., & Clarke, V. (2006). Using thematic analysis in psychology. Qualitative Research in Psychology, 3(2), 77–101. https://doi.org/10.1191/1478088706qp063oa

Bylund, C. L., & Makoul, G. (2002). Empathic communication and gender in the physician–patient encounter. Patient education and counseling, 48(3), 207-216. https://doi.org/10.1016/S0738-3991(02)00173-8

Decety, J., & Cowell, J. M. (2014). The complex relation between morality and empathy. Trends in cognitive sciences, 18(7), 337-339. https://doi.org/10.1016/j.tics.2018.02.003

Dennett, D. C. (1987). The intentional stance. MIT Press.

Epley, N., Waytz, A., & Cacioppo, J. T. (2007). On seeing human: A three-factor theory of anthropomorphism. Psychological Review, 114(4), 864–886. https://doi.org/10.1037/0033-295X.114.4.864

Li, Y., Zhang, W., & Xu, H. (2024). Are large language models more empathetic than humans? arXiv Preprint. https://arxiv.org/abs/2406.05063

Ma, N., Wang, X., & Chen, J. (2025). Effect of anthropomorphism and perceived intelligence in conversational agents. Frontiers in Computer Science, 7, 1531976. https://doi.org/10.3389/fcomp.2025.1531976

Mercer, S. W., Maxwell, M., Heaney, D., & Watt, G. C. (2004). The consultation and relational empathy (CARE) measure: Development and preliminary validation. Family Practice, 21(6), 699–705. https://doi.org/10.1093/fampra/cmh621

Nair, R., Gupta, S., & Thomas, P. (2025). Large language models for surgical informed consent: Empathy, ethics, and efficacy. Journal of Medical Ethics. https://doi.org/10.1136/jme-2024-110652

Pei, H., Yang, F., & Zhou, Z. (2024). Affective computing: Recent advances, challenges, and future trends. Frontiers in Computer Science, 6, 1548389. https://doi.org/10.34133/icomputing.0076

Picard, R. W. (1997). Affective computing. MIT Press.

Richet, J. L. (2025). AI companionship or digital entrapment? Investigating the impact of anthropomorphic AI-based chatbots. Journal of Innovation & Knowledge, 10(6), 100835. https://doi.org/10.1016/j.jik.2025.100835

Schlegel, K., Schmid Mast, M., & García, D. (2025). Large language models are proficient in solving and generating emotional intelligence tests. npj Science of Learning, 10(1), 1–9. https://doi.org/10.1038/s44271-025-00258-x

Searle, J. R. (1980). Minds, brains, and programs. Behavioral and Brain Sciences, 3(3), 417–457. https://doi.org/10.1017/S0140525X00005756

Shamay-Tsoory, S. G. (2011). The neural bases for empathy. The Neuroscientist, 17(1), 18–24. https://doi.org/10.1177/1073858410379268

Shen, J., DiPaola, D., Ali, S., Sap, M., Park, H., & Breazeal, C. (2024). Empathy toward AI versus human experiences in mental health chatbot design. JMIR Mental Health, 11, e62679. https://doi.org/10.2196/62679

Sorin, V., Brin, D., Barash, Y., Konen, E., Charney, A., Nadkarni, G., & Klang, E. (2024). Large language models and empathy: Systematic review. Journal of Medical Internet Research, 26, e52597. https://doi.org/10.2196/52597

Sung, L. U. (2025). Empathetic large language models and the meaning of empathy. Inquiry. https://doi.org/10.1080/0020174X.2025.2518448

Truong, T. T. H., & Chen, J. S. (2025). When empathy is enhanced by human–AI interaction: An investigation of anthropomorphism and responsiveness on customer experience with AI chatbots. Asia Pacific Journal of Marketing and Logistics, 37(12), 3908-3925. https://doi.org/10.1108/APJML-10-2024-1464

Weizenbaum, J., & McCarthy, J. (1977). Computer power and human reason: From judgment to calculation.

Zaki, J. (2019). The war for kindness: Building empathy in a fractured world. Crown.

Zhang, Y., Yang, X., Xu, X., Gao, Z., Huang, Y., Mu, S., ... & Yu, G. (2026). Affective computing in the era of large language models: A survey from the nlp perspective. Knowledge-Based Systems, 115411. https://doi.org/10.1016/j.knosys.2026.115411

Downloads

Published

2026-06-29

Issue

Section

Vol. 13, No. 1 (2026)

How to Cite

Mugableh, A., & Mohammed Alruwaili, H. (2026). Functional vs. phenomenological empathy in large language models: Rethinking artificial empathy through experimental evidence. East European Journal of Psycholinguistics , 13(1), 307-332. https://doi.org/10.29038/eejpl.2026.13.1.mug

Similar Articles

1-10 of 288

You may also start an advanced similarity search for this article.