• Home
  • Search
  • Lexicon-Based Graph Convolutional Network for Chinese Word Segmentation
  • https://doi.org/10.48448/m2bs-bb02Copy DOI Icon

Lexicon-Based Graph Convolutional Network for Chinese Word Segmentation

Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

Precise information of word boundary can alleviate the problem of lexical ambiguity to improve the performance of natural language processing (NLP) tasks. Thus, Chinese word segmentation (CWS) is a fundamental task in NLP. Due to the development of pre-trained language models (PLM), pre-trained knowledge can help neural methods solve the main problems of the CWS in significant measure. Existing methods have already achieved high performance on several benchmarks (e.g., Bakeoff-2005). However, recent outstanding studies are limited by the small-scale annotated corpus. To further improve the performance of CWS methods based on fine-tuning the PLMs, we propose a novel neural framework, LBGCN, which incorporates a lexicon-based graph convolutional network into the Transformer encoder. Experimental results on five benchmarks and four cross-domain datasets show the lexicon-based graph convolutional network successfully captures the information of candidate words and helps to improve performance on the benchmarks (Bakeoff-2005 and CTB6) and the cross-domain datasets (SIGHAN-2010). Further experiments and analyses demonstrate that our proposed framework effectively models the lexicon to enhance the ability of basic neural frameworks and strengthens the robustness in the cross-domain scenario.

Similar Papers
  • Video Transcripts

Can Pre-trained Language Models Interpret Similes as Smart as Human?

  • May 11, 2022
  • Underline Science Inc.
  • Qianyu He +4
  • PDF
  • Research Article
  • Citations7

APRE: Annotation-Aware Prompt-Tuning for Relation Extraction

  • Feb 21, 2024
  • Neural Processing Letters
  • Chao Wei +5
  • Research Article
  • Citations20

A Survey on Automatic Generation of Figurative Language: From Rule-based Systems to Large Language Models

  • May 14, 2024
  • ACM Computing Surveys
  • Huiyuan Lai +1
  • Research Article
  • Citations6

Deep Fusing Pre-trained Models into Neural Machine Translation

  • Jun 28, 2022
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Rongxiang Weng +3
  • Research Article
  • Citations8

Point-to-Pixel Prompting for Point Cloud Analysis With Pre-Trained Image Models.

  • Jun 01, 2024
  • IEEE Transactions on Pattern Analysis and Machine Intelligence
  • Ziyi Wang +4
  • PDF
  • Research Article
  • Citations35

Unrestricted Attention May Not Be All You Need–Masked Attention Mechanism Focuses Better on Relevant Parts in Aspect-Based Sentiment Analysis

  • Jan 01, 2022
  • IEEE Access
  • Ao Feng +2
  • Research Article

TOWARDS CROSS-ATTENTION PRE-TRAINING IN NEURAL MACHINE TRANSLATION

  • Oct 31, 2022
  • Tạp chí Khoa học
  • Khang Pham
  • PDF
  • Research Article
  • Citations16

BioBERTurk: Exploring Turkish Biomedical Language Model Development Strategies in Low-Resource Setting.

  • Sep 19, 2023
  • Journal of healthcare informatics research
  • Hazal Türkmen +4
  • Research Article

Task-Adaptive and Multi-Level Contextual Understanding for Emotion Recognition in Conversations

  • Feb 09, 2026
  • Applied Sciences
  • Xiaomeng Yao +4
  • Research Article
  • Citations3

Exploring Named Entity Recognition via MacBERT-BiGRU and Global Pointer with Self-Attention

  • Dec 03, 2024
  • Big Data and Cognitive Computing
  • Chengzhe Yuan +6
  • Research Article
  • Citations20

Prompt for extraction: Multiple templates choice model for event extraction

  • Feb 20, 2024
  • Knowledge-Based Systems
  • Jiaren Peng +3
  • Research Article
  • Citations1

SensiMix: Sensitivity-Aware 8-bit index & 1-bit value mixed precision quantization for BERT compression

  • Apr 18, 2022
  • PLoS ONE
  • Tairen Piao +3
  • Dissertation
  • Citations1

Natural language processing as autoregressive generation

  • Jan 01, 2023
  • Xiang Lin
  • PDF
  • Conference Article
  • Citations22

Causal-Debias: Unifying Debiasing in Pretrained Language Models and Fine-tuning via Causal Invariant Learning

  • Jan 01, 2023
  • Fan Zhou +4
  • PDF
  • Research Article
  • Citations89

From Word Embeddings to Pre-Trained Language Models: A State-of-the-Art Walkthrough

  • Sep 01, 2022
  • Applied Sciences
  • Mourad Mars
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.