Symbolic Faces and Artificial Minds: Evaluating Artificial Intelligence Recognition of Punctuation-Based Face Expressions Using ChatGPT, Claude, and DeepSeek Models

Baker RSJd, Siemens G. Educational data mining and learning analytics. In: Sawyer K, editor. Cambridge handbook of the learning sciences (2nd ed.). 2014.

Holmes W, Bialik M, Fadel C. Artificial intelligence in education: promises and implications for teaching and learning. Center for Curriculum Redesign. Boston (MA). 2019.

Obermeyer Z, Powers B, Vogeli C, Mullainathan S. Dissecting racial bias in an algorithm used to manage the health of populations. Science. 2019;366(6464):447–53.

Article  Google Scholar 

Jiang F, Jiang Y, Zhi H, Dong Y, Li H, Ma S, Wang Y, Dong Q, Shen H, Wang Y. Artificial intelligence in healthcare: past, present and future. Stroke Vasc Neurol. 2017. https://doi.org/10.1136/svn-2017-000101.

Article  Google Scholar 

Roumeliotis KI, Tselikas ND. ChatGPT and Open-AI models: a preliminary review. Future Internet. 2023;15(6): 192.

Article  Google Scholar 

Kundu S, Bai Y, Kadavath S, Askell A, Callahan A, Chen A, Goldie A, Balwit A, Mirhoseini A, McLean B, Olsson C. Specific versus general principles for constitutional ai. arXiv preprint arXiv:2310.13798. 2023. https://doi.org/10.48550/arXiv.2310.13798.

Peng Y, Malin BA, Rousseau JF, Wang Y, Xu Z, Xu X, Weng C, Bian J. From GPT to DeepSeek: significant gaps remain in realizing AI in healthcare. J Biomed Inform. 2025;163: 104791.

Article  Google Scholar 

Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, Agarwal S. Language models are few-shot learners. Adv Neural Inf Process Syst. 2020;33:1877–901.

Google Scholar 

Liu A, Feng B, Xue B, Wang B, Wu B, Lu C, Zhao C, Deng C, Zhang C, Ruan C, Dai D. Deepseek-v3 technical report. arXiv preprint 2024. arXiv:2412.19437.

Wang C, Szarvas G, Balazs G, Danchenko P, Ernst P. Calibrating Verbalized Probabilities for Large Language Models. arXiv preprint 2024 arXiv:2410.06707.

Tomić MK, Marinez M, Vrbanec T. Emoticons. FIP - J Finance Law Effectus. 2014;1:35–42.

Google Scholar 

Bai Q, Dan Q, Mu Z, et al. A systematic review of emoji: current research and future perspectives. Front Psychol. 2019;10: 2221.

Article  Google Scholar 

Gantiva C, Sotaquirá M, Araujo A, et al. Cortical processing of human and emoji faces: an ERP analysis. Behav Inf Technol. 2020;39:935–43.

Article  Google Scholar 

Keskin EK. Yapay Zekâ Sohbet Robotu ChatGPT ve Türkiye İnternet Gündeminde Oluşturduğu Temalar (In Turkish). Yeni Medya Elektronik Dergisi. 2023;7(2):114–31.

Costanza R, De Groot R, Sutton P, et al. Changes in the global value of ecosystem services. Glob Environ Chang. 2014;26:152–8.

Article  Google Scholar 

Verpoorter C, Kutser T, Seekell DA, et al. A global inventory of lakes based on high-resolution satellite imagery. Geophys Res Lett. 2014;41:6396–402.

Article  Google Scholar 

Messager ML, Lehner B, Grill G, et al. Estimating the volume and age of water stored in global lakes using a geo-statistical approach. Nat Commun. 2016;7(1): 13603.

Article  Google Scholar 

Tsigaris P, Teixeira da Silva JA. Can ChatGPT be trusted to provide reliable estimates?. Account Res. 2023;31(7):1–3.

Avnat E, et al. Performance of large language models in numerical vs. semantic medical knowledge: benchmarking on evidence-based QandAs. arXiv preprint arXiv:2406.03855. 2024.

Gupta GK, Pande P. LLMs in disease diagnosis: a comparative study of DeepSeek-R1 and O3 mini across chronic health conditions. arXiv preprint arXiv:2503.10486. 2025.

Štular B, Lozić E. ChatGPT v Bard v Bing v Claude 2 v Aria v human-expert. How good are AI chatbots at scientific writing? arXiv:2309.08636. 2023.

Uppalapati VK, Nag DS. A comparative analysis of AI models in complex medical decision-making scenarios: evaluating ChatGPT, Claude AI, Bard, and Perplexity. Cureus. 2024;16(1):e52485. https://doi.org/10.7759/cureus.52485.

Albuhairy MM, Algaraady J. DeepSeek vs. ChatGPT: comparative efficacy in reasoning for adults’ second language acquisition analysis. Center for Open Science. 2025.

Kotsis KT. ChatGPT and DeepSeek evaluate one another for science education. EIKI J Eff Teach Methods. 2025;3(1):98–102. https://doi.org/10.59652/jetm.v3i1.439.

Mondillo G, Colosimo S, Frattolillo V, Perrotta A, Masino M. Comparative evaluation of advanced AI reasoning models in pediatric clinical decision support: ChatGPT O1 vs. DeepSeek-R1. Cold Spring Harbor Laboratory. 2025.

Adolphs R. Recognizing emotion from facial expressions: psychological and neurological mechanisms. Behav Cogn Neurosci Rev. 2002;1(1):21–62. https://doi.org/10.1177/1534582302001001003.

Barrett LF, Mesquita B, Gendron M. Context in emotion perception. Curr Dir Psychol Sci. 2011;20(5):286–90.

Article  Google Scholar 

Haxby JV, Hoffman EA, Gobbini MI. The distributed human neural system for face perception. Trends Cogn Sci. 2000;4(6):223–33.

Article  Google Scholar 

Kanwisher N, McDermott J, Chun MM. The fusiform face area: a module in human extrastriate cortex specialized for face perception. J Neurosci. 1997;17(11):4302–11.

Article  Google Scholar 

Pitcher D, Walsh V, Yovel G, Duchaine B. TMS evidence for the involvement of the right occipital face area in early face processing. Curr Biol. 2008;17(18):1568–73.

Article  Google Scholar 

Zhen Z, Fang H, Liu J. The hierarchical brain network for face recognition. PLoS One. 2013;8: e59886.

Article  Google Scholar 

Haxby JV, Gobbini MI, Furey ML, Ishai A, Schouten JL, Pietrini P. Distributed and overlapping representations of faces and objects in ventral temporal cortex. Science. 2001;293(5539):2425–30.

Article  Google Scholar 

Bernstein M, Yovel G. Two neural pathways of face processing: a critical evaluation of current models. Neurosci Biobehav Rev. 2015;55:536–46.

Article  Google Scholar 

Churches O, Nicholls M, Thiessen M, et al. Emoticons in mind: an event-related potential study. Soc Neurosci. 2014;9(2):196–202.

Article  Google Scholar 

Cao J, Zhao L. The electrophysiological correlates of internet language processing revealed by N170 elicited by network emoticons. Neuroreport. 2018;29:1055–60.

Article  Google Scholar 

Akdeniz G. Brain activity underlying face and face pareidolia processing: an ERP study. Neurol Sci. 2020;41:1557–65.

Article  Google Scholar 

Ganel T, Sofer C, Goodale MA. Biases in human perception of facial age are present and more exaggerated in current AI technology. Sci Rep. 2022;12:22519.

Article  Google Scholar 

Ouyang L, Wu J, Jiang X, Almeida D, Wainwright C, Mishkin P, et al. Training language models to follow instructions with human feedback. arXiv preprint. 2022 arXiv:2203.02155.

Zhou M, Ding M, Zhou J. Evaluating the minimal context comprehension ability of large language models. arXiv preprint arXiv:2305.13267. 2023.

Bai Y, Kadavath S, Kundu S, Askell A, Kernion J, Jones A, Chen A, Goldie A, Mirhoseini A, McKinnon C, Chen C. Constitutional AI: harmlessness from AI feedback. 2022. arXiv preprint arXiv:2212.08073 . 2022.

Johnson D, Goodman R, Patrinely J, et al. Assessing the accuracy and reliability of AI-generated medical responses. An Evaluation of the Chat-GPT Model. Res Sq. 2023;rs.3.rs-2566942. https://doi.org/10.21203/rs.3.rs-2566942/v1.

Comments (0)

No login
gif