• Home
  • Search
  • Pre-trained language models with domain knowledge for biomedical extractive summarization
  • Cite Icon87
  • https://doi.org/10.1016/j.knosys.2022.109460Copy DOI Icon

Pre-trained language models with domain knowledge for biomedical extractive summarization

Show More
  • Abstract
  • Highlights & Summary
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Biomedical text summarization is a critical task for comprehension of an ever-growing amount of biomedical literature. Pre-trained language models (PLMs) with transformer-based architectures have been shown to greatly improve performance in biomedical text mining tasks. However, existing methods for text summarization generally fine-tune PLMs on the target corpora directly and do not consider how fine-grained domain knowledge, such as PICO elements used in evidence-based medicine, can help to identify the context needed for generating coherent summaries. To fill the gap, we propose KeBioSum, a novel knowledge infusion training framework, and experiment using a number of PLMs as bases, for the task of extractive summarization on biomedical literature. We investigate generative and discriminative training techniques to fuse domain knowledge (i.e., PICO elements) into knowledge adapters and apply adapter fusion to efficiently inject the knowledge adapters into the basic PLMs for fine-tuning the extractive summarization task. Experimental results from the extractive summarization task on three biomedical literature datasets show that existing PLMs (BERT, RoBERTa, BioBERT, and PubMedBERT) are improved by incorporating the KeBioSum knowledge adapters, and our model outperforms the strong baselines.

Similar Papers
  • PDF
  • Research Article
  • Citations10

AGI-P: A Gender Identification Framework for Authorship Analysis Using Customized Fine-Tuning of Multilingual Language Model

  • Jan 01, 2024
  • IEEE Access
  • Raheem Sarwar +6
  • Conference Article
  • Citations3

Chinese-Korean Weibo Sentiment Classification Based on Pre-trained Language Model and Transfer Learning

  • May 06, 2022
  • Hengxuan Wang +3
  • Research Article
  • Citations2

Large Language Models Evaluation for PubMed Extractive Summarisation

  • Jan 14, 2026
  • ACM Transactions on Computing for Healthcare
  • Tian Cheng Xia +2
  • Research Article

A method to utilize prior knowledge for extractive summarization based on pre-trained language models

  • Dec 05, 2024
  • Vietnam Journal of Science and Technology
  • Le Ngoc Thang +4
  • PDF
  • Conference Article
  • Citations31

PANLP at MEDIQA 2019: Pre-trained Language Models, Transfer Learning and Knowledge Distillation

  • Jan 01, 2019
  • Wei Zhu +6
  • Video Transcripts

Domain Knowledge Transferring for Pre-trained Language Model via Calibrated Activation Boundary Distillation

  • May 07, 2022
  • Underline Science Inc.
  • Dongha Choi +2
  • Research Article
  • Citations20

Entity recognition in the field of coal mine construction safety based on a pre-training language model

  • Dec 28, 2023
  • Engineering, Construction and Architectural Management
  • Na Xu +6
  • Research Article
  • Citations13

JointMatcher: Numerically-aware entity matching using pre-trained language models with attention concentration

  • May 16, 2022
  • Knowledge-Based Systems
  • Chen Ye +6
  • Book Chapter
  • Citations5

Extractive Summarization of Chinese Judgment Documents via Sentence Embedding and Memory Network

  • Jan 01, 2021
  • Yan Gao +3
  • Research Article
  • Citations59

A systematic review of automatic text summarization for biomedical literature and EHRs.

  • Aug 02, 2021
  • Journal of the American Medical Informatics Association
  • Mengqian Wang +5
  • Conference Article
  • Citations4

StyleBERT: Chinese Pretraining by Font Style Information

  • Jun 17, 2022
  • Chao Lv +7
  • PDF
  • Research Article
  • Citations3

Tibetan Sentence Boundaries Automatic Disambiguation Based on Bidirectional Encoder Representations from Transformers on Byte Pair Encoding Word Cutting Method

  • Apr 02, 2024
  • Applied Sciences
  • Fenfang Li +3
  • Conference Article

Exploring Layer-wise Representations of English and Chinese Homonymy in Pre-trained Language Models

  • Jan 01, 2025
  • Matthew King-Hang Ma +3
  • Research Article
  • Citations10

UniproLcad: Accurate Identification of Antimicrobial Peptide by Fusing Multiple Pre-Trained Protein Language Models

  • Apr 11, 2024
  • Symmetry
  • Xiao Wang +3
  • PDF
  • Research Article
  • Citations10

Classification and analysis of text transcription from Thai depression assessment tasks among patients with depression.

  • Mar 30, 2023
  • PLOS ONE
  • Adirek Munthuli +8
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.