• Home
  • Search
  • Universal Sentence Representation Learning with Conditional Masked Language Model
  • Cite Icon39
  • https://doi.org/10.18653/v1/2021.emnlp-main.502Copy DOI Icon

Universal Sentence Representation Learning with Conditional Masked Language Model

  • Jan 1, 2021
  • Ziyi Yang +4 more
Show More
  • Abstract
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This paper presents a novel training method, Conditional Masked Language Modeling (CMLM), to effectively learn sentence representations on large scale unlabeled corpora. CMLM integrates sentence representation learning into MLM training by conditioning on the encoded vectors of adjacent sentences. Our English CMLM model achieves state-ofthe-art performance on SentEval (Conneau and Kiela, 2018), even outperforming models learned using supervised signals. As a fully unsupervised learning method, CMLM can be conveniently extended to a broad range of languages and domains. We find that a multilingual CMLM model co-trained with bitext retrieval (BR) and natural language inference (NLI) tasks outperforms the previous state-of-the-art multilingual models by a large margin, e.g. 10% improvement upon baseline models on cross-lingual semantic search. We explore the same language bias of the learned representations, and propose a simple, post-training and model agnostic approach to remove the language identifying information from the representation while still retaining sentence semantics.

Loading PDF

Similar Papers
  • Research Article
  • Citations12

Usr-mtl: an unsupervised sentence representation learning framework with multi-task learning

  • Nov 14, 2020
  • Applied Intelligence
  • Wenshen Xu +2
  • Conference Article
  • Citations4

Resolving Word Vagueness with Scenario-guided Adapter for Natural Language Inference

  • Aug 01, 2024
  • Xinrui Lin +6
  • Conference Article

Simple Temperature Cool-down in Contrastive Framework for Unsupervised Sentence Representation Learning

  • Jan 01, 2024
  • Yoo Hyun Jeong +2
  • Book Chapter
  • Citations6

Training Dataset and Dictionary Sizes Matter in BERT Models: The Case of Baltic Languages

  • Jan 01, 2022
  • Matej Ulčar +1
  • PDF
  • Research Article

UniBERT: adversarial training for language-universal representations

  • Aug 04, 2025
  • Neural Computing and Applications
  • Andrei-Marius Avram +4
  • Video Transcripts

Learning Natural Language Generation with Truncated Reinforcement Learning

  • Jul 10, 2022
  • Underline Science Inc.
  • Alice Martin
  • Video Transcripts

Embracing Ambiguity: Shifting the Training Target of NLI Models

  • Aug 01, 2021
  • Underline Science Inc.
  • Johannes Mario Meissner +3
  • Research Article
  • Citations6

Natural language inference for Malayalam language using language agnostic sentence representation

  • May 04, 2021
  • PeerJ Computer Science
  • Sara Renjit +1
  • PDF
  • Conference Article
  • Citations10

ERNIE-NLI: Analyzing the Impact of Domain-Specific External Knowledge on Enhanced Representations for NLI

  • Jan 01, 2021
  • Lisa Bauer +2
  • PDF
  • Research Article

Extracting Reproductive Condition and Habitat Information from Text Using a Transformer-based Information Extraction Pipeline

  • Sep 11, 2023
  • Biodiversity Information Science and Standards
  • Roselyn Gabud +3
  • Book Chapter
  • Citations1

Natural Language Inference Using Evidence from Knowledge Graphs

  • Jan 01, 2021
  • Boxuan Jia +2
  • Research Article

ViMMRC 2.0 — Enhancing Machine Reading Comprehension on Vietnamese Literature Text

  • Jul 15, 2025
  • International Journal of Asian Language Processing
  • Son T Luu +4
  • Video Transcripts

Towards Debiasing Translation Artifacts

  • Jun 27, 2022
  • Underline Science Inc.
  • Koel Dutta Chowdhury
  • Research Article

Exploring Selective Layer Freezing Strategies in Transformer Fine-Tuning: NLI Classifiers with Sub-3B Parameter Models

  • Sep 26, 2025
  • Applied Sciences
  • Taewook Hwang +3
  • Conference Article

Unsupervised Sentence Representation Learning with Syntactically Aligned Negative Samples

  • Jan 01, 2025
  • Zhilan Wang +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.