• Home
  • Search
  • Does GPT-4 surpass human performance in linguistic pragmatics?
  • Cite Icon10
  • https://doi.org/10.1057/s41599-025-04912-xCopy DOI Icon

Does GPT-4 surpass human performance in linguistic pragmatics?

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

As Large Language Models (LLMs) become increasingly integrated into everyday life as general-purpose multimodal AI systems, their capabilities to simulate human understanding are under examination. This study investigates LLMs’ ability to interpret linguistic pragmatics, which involves context and implied meanings. Using Grice’s communication principles, we evaluated both LLMs (GPT-2, GPT-3, GPT-3.5, GPT-4, and Bard) and human subjects (N = 147) on dialogue-based tasks. Human participants included 71 primarily Serbian students and 76 native English speakers from the United States. Findings revealed that LLMs, particularly GPT-4, outperformed humans. GPT-4 achieved the highest score of 4.80, surpassing the best human score of 4.55. Other LLMs performed well: GPT-3.5 scored 4.10, Bard 3.75, and GPT-3 3.25; GPT-2 had the lowest score of 1.05. The average LLM score was 3.39, exceeding the human cohorts’ averages of 2.80 (Serbian students) and 2.34 (U.S. participants). In the ranking of all 155 subjects (including LLMs and humans), GPT-4 secured the top position, while the best human ranked second. These results highlight significant progress in LLMs’ ability to simulate understanding of linguistic pragmatics. Future studies should confirm these findings with more dialogue-based tasks and diverse participants. This research has important implications for advancing general-purpose AI models in various communication-centered tasks, including potential application in humanoid robots in the future.

Similar Papers
  • Research Article

Research and selection of Large Learning Models for automation of ABAP-code migration

  • Sep 24, 2025
  • Management of Development of Complex Systems
  • Oleg Pozdnyakov +1
  • Research Article

Large language model use in oral and maxillofacial surgery training: a national resident survey.

  • Feb 21, 2026
  • Oral and maxillofacial surgery
  • Nolan Kranc +7
  • Research Article
  • Citations3

Decoding Multilingual Moral Preferences: Unveiling LLM's Biases through the Moral Machine Experiment

  • Oct 16, 2024
  • Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society
  • Karina Vida +2
  • Research Article
  • Citations14

Leveraging large language models for comprehensive locomotion control in humanoid robots design

  • Oct 16, 2024
  • Biomimetic Intelligence and Robotics
  • Shilong Sun +4
  • Research Article

A Multimodal AI System: Comparing LLMs and Theorem Proving Systems

  • Feb 21, 2026
  • Electronics
  • Phillip G Bradford +1
  • Research Article

Discrete Tokenization for Multimodal LLMs: A Comprehensive Survey.

  • Mar 24, 2026
  • IEEE transactions on pattern analysis and machine intelligence
  • Jindong Li +7
  • Research Article

Full-Stack Optimized Large Language Models for Lifelong Sequential Behavior Comprehension in Recommendation

  • Nov 21, 2025
  • ACM Transactions on Recommender Systems
  • Rong Shan +7
  • Research Article
  • Citations18

DracoGPT: Extracting Visualization Design Preferences from Large Language Models.

  • Jan 01, 2025
  • IEEE transactions on visualization and computer graphics
  • Huichen Will Wang +3
  • Research Article
  • Citations2

Can general purpose large language models assist pediatricians in predicting infants with serious bacterial infection?

  • Nov 14, 2025
  • BMC Medical Informatics and Decision Making
  • Ivan Šimunović +9
  • Conference Article

Data-Efficient Tabular Classification with Transformer-Based Small Language Models

  • Nov 10, 2025
  • Mario Haddad-Neto +3
  • Conference Article

LinkGPT: Leveraging Large Language Models for Enhanced Link Prediction in Text-Attributed Graphs

  • Nov 07, 2025
  • Zhongmou He +4
  • Research Article
  • Citations5

Medical Students' Perceptions of Large Language Models in Healthcare: A Multinational Cross-Sectional Study.

  • May 21, 2025
  • Journal of medical education and curricular development
  • Faiza Ejas +22
  • Preprint Article

It Knew Too Much: On the Unsuitability of LLMs as Replacements for Human Subjects

  • May 28, 2025
  • Amaç Herdağdelen +1
  • Research Article

Generation of machine-readable country-by-country reports with large language models

  • Aug 29, 2025
  • Eastern-European Journal of Enterprise Technologies
  • Yakiv Yusyn
  • PDF
  • Research Article
  • Citations3

Exploring the potential of lightweight large language models for AI-based mental health counselling task: a novel comparative study

  • Jul 02, 2025
  • Scientific Reports
  • Ritesh Maurya +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.