• Home
  • Search
  • A method to utilize prior knowledge for extractive summarization based on pre-trained language models
  • https://doi.org/10.15625/2525-2518/20241Copy DOI Icon

A method to utilize prior knowledge for extractive summarization based on pre-trained language models

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

This paper presents a novel model for extractive summarization that integrates context representation from a pre-trained language model (PLM), such as BERT, with prior knowledge derived from unsupervised learning methods. Sentence importance assessment is crucial in extractive summarization, with prior knowledge providing indicators of sentence importance within a document. Our model introduces a method for estimating sentence importance based on prior knowledge, complementing the contextual representation offered by PLMs like BERT. Unlike previous approaches that primarily relied on PLMs alone, our model leverages both contextual representation and prior knowledge extracted from each input document. By conditioning the model on prior knowledge, it emphasizes key sentences in generating the final summary. We evaluate our model on three benchmark datasets across two languages, demonstrating improved performance compared to strong baseline methods in extractive summarization. Additionally, our ablation study reveals that injecting knowledge into certain first attention layers yields greater benefits than others. The model code is publicly available for further exploration.

Similar Papers
  • Research Article
  • Citations87

Pre-trained language models with domain knowledge for biomedical extractive summarization

  • Jul 19, 2022
  • Knowledge-Based Systems
  • Qianqian Xie +3
  • PDF
  • Research Article
  • Citations10

AGI-P: A Gender Identification Framework for Authorship Analysis Using Customized Fine-Tuning of Multilingual Language Model

  • Jan 01, 2024
  • IEEE Access
  • Raheem Sarwar +6
  • Conference Article
  • Citations3

Chinese-Korean Weibo Sentiment Classification Based on Pre-trained Language Model and Transfer Learning

  • May 06, 2022
  • Hengxuan Wang +3
  • PDF
  • Conference Article
  • Citations31

PANLP at MEDIQA 2019: Pre-trained Language Models, Transfer Learning and Knowledge Distillation

  • Jan 01, 2019
  • Wei Zhu +6
  • Conference Article
  • Citations4

StyleBERT: Chinese Pretraining by Font Style Information

  • Jun 17, 2022
  • Chao Lv +7
  • PDF
  • Research Article
  • Citations3

Tibetan Sentence Boundaries Automatic Disambiguation Based on Bidirectional Encoder Representations from Transformers on Byte Pair Encoding Word Cutting Method

  • Apr 02, 2024
  • Applied Sciences
  • Fenfang Li +3
  • Research Article
  • Citations4

Channel-Aware Decoupling Network for Multiturn Dialog Comprehension.

  • Jun 01, 2024
  • IEEE transactions on neural networks and learning systems
  • Zhuosheng Zhang +2
  • Research Article
  • Citations13

JointMatcher: Numerically-aware entity matching using pre-trained language models with attention concentration

  • May 16, 2022
  • Knowledge-Based Systems
  • Chen Ye +6
  • Conference Article

Exploring Layer-wise Representations of English and Chinese Homonymy in Pre-trained Language Models

  • Jan 01, 2025
  • Matthew King-Hang Ma +3
  • Research Article
  • Citations10

UniproLcad: Accurate Identification of Antimicrobial Peptide by Fusing Multiple Pre-Trained Protein Language Models

  • Apr 11, 2024
  • Symmetry
  • Xiao Wang +3
  • PDF
  • Research Article
  • Citations10

Classification and analysis of text transcription from Thai depression assessment tasks among patients with depression.

  • Mar 30, 2023
  • PLOS ONE
  • Adirek Munthuli +8
  • Conference Article
  • Citations16

Unsupervised Neural Machine Translation for English to Kannada Using Pre-Trained Language Model

  • Oct 03, 2022
  • Shailashree K Sheshadri +4
  • Research Article
  • Citations1

Using a large language model to provide individualized feedback for pre-service physics teachers’ written reflections

  • Nov 21, 2025
  • Disciplinary and Interdisciplinary Science Education Research
  • Stefan Sorge +2
  • Video Transcripts

Robust Transfer Learning with Pretrained Language Models through Adapters

  • Aug 01, 2021
  • Underline Science Inc.
  • Wenjuan Han +2
  • Video Transcripts

Can Pre-trained Language Models Interpret Similes as Smart as Human?

  • May 11, 2022
  • Underline Science Inc.
  • Qianyu He +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.