• Home
  • Search
  • Explainability for Large Language Models: A Survey
  • Cite Icon503
  • https://doi.org/10.1145/3639372Copy DOI Icon

Explainability for Large Language Models: A Survey

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Large language models (LLMs) have demonstrated impressive capabilities in natural language processing. However, their internal mechanisms are still unclear and this lack of transparency poses unwanted risks for downstream applications. Therefore, understanding and explaining these models is crucial for elucidating their behaviors, limitations, and social impacts. In this article, we introduce a taxonomy of explainability techniques and provide a structured overview of methods for explaining Transformer-based language models. We categorize techniques based on the training paradigms of LLMs: traditional fine-tuning-based paradigm and prompting-based paradigm. For each paradigm, we summarize the goals and dominant approaches for generating local explanations of individual predictions and global explanations of overall model knowledge. We also discuss metrics for evaluating generated explanations and discuss how explanations can be leveraged to debug models and improve performance. Lastly, we examine key challenges and emerging opportunities for explanation techniques in the era of LLMs in comparison to conventional deep learning models.

Loading PDF

Similar Papers
  • Research Article
  • Citations51

Large language models for biomedicine: foundations, opportunities, challenges, and best practices.

  • Apr 24, 2024
  • Journal of the American Medical Informatics Association : JAMIA
  • Satya S Sahoo +8
  • Research Article

#2924 Comparison of large language models and traditional natural language processing techniques in predicting arteriovenous fistula failure

  • May 23, 2024
  • Nephrology Dialysis Transplantation
  • Suman Lama +6
  • Research Article

Evaluating gpt-4 for zero-shot classification of bleeding and clotting events: Can large language models serve as second reviewers?

  • Nov 03, 2025
  • Blood
  • Samantha Rizzo +5
  • Conference Article

Data-Efficient Tabular Classification with Transformer-Based Small Language Models

  • Nov 10, 2025
  • Mario Haddad-Neto +3
  • Research Article
  • Citations7

The Generalization and Robustness of Transformer-Based Language Models on Commonsense Reasoning

  • Mar 24, 2024
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Ke Shen
  • PDF
  • Research Article
  • Citations13

Advancements and Applications of Large Language Models in Natural Language Processing: A Comprehensive Review

  • Nov 26, 2024
  • Applied and Computational Engineering
  • Mengchao Ren
  • Research Article
  • Citations14

Fine-Tuned Understanding: Enhancing Social Bot Detection With Transformer-Based Classification

  • Jan 01, 2024
  • IEEE Access
  • Amine Sallah +6
  • Supplementary Content

Large language and vision-language models for robot: safety challenges, mitigation strategies and future directions

  • Jul 29, 2025
  • Industrial Robot: the international journal of robotics research and application
  • Xiangyu Hu +1
  • Research Article
  • Citations2

Toward Cross-Hospital Deployment of Natural Language Processing Systems: Model Development and Validation of Fine-Tuned Large Language Models for Disease Name Recognition in Japanese

  • Jul 08, 2025
  • JMIR Medical Informatics
  • Seiji Shimizu +4
  • Research Article
  • Citations1

Urdu Sentential Paraphrased Plagiarism Detection Using Large Language Models

  • Jul 11, 2025
  • ACM Transactions on Asian and Low-Resource Language Information Processing
  • Hafiz Rizwan Iqbal +4
  • Research Article
  • Citations1

Transforming scholarly landscapes: The influence of large language models on academic fields beyond computer science.

  • Jan 14, 2026
  • PloS one
  • Aniket Pramanick +3
  • Research Article
  • Citations3

Automated Extraction of Mortality Information From Publicly Available Sources Using Large Language Models: Development and Evaluation Study

  • Aug 18, 2025
  • Journal of Medical Internet Research
  • Mohammed Al-Garadi +13
  • Research Article

Transformative Trends: A Comprehensive Review of Large Language Models (LLMs) in Healthcare

  • Jun 02, 2024
  • INTERANTIONAL JOURNAL OF SCIENTIFIC RESEARCH IN ENGINEERING AND MANAGEMENT
  • Chetna Kumari
  • Research Article
  • Citations1

Assessment of ChatGPT performance in orbital MRI reporting with multimetric evaluation of transformer based language models

  • Oct 13, 2025
  • Scientific Reports
  • Alessandro Tel +4
  • Research Article

Research and selection of Large Learning Models for automation of ABAP-code migration

  • Sep 24, 2025
  • Management of Development of Complex Systems
  • Oleg Pozdnyakov +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.