• Home
  • Search
  • Czert – Czech BERT-like Model for Language Representation
  • Cite Icon21
  • https://doi.org/10.26615/978-954-452-072-4_149Copy DOI Icon

Czert – Czech BERT-like Model for Language Representation

  • Jan 1, 2021
  • Jakub Sido +5 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

This paper describes the training process of the first Czech monolingual language representation models based on BERT and ALBERT architectures. We pre-train our models on more than 340K of sentences, which is 50 times more than multilingual models that include Czech data. We outperform the multilingual models on 9 out of 11 datasets. In addition, we establish the new state-of-the-art results on nine datasets. At the end, we discuss properties of monolingual and multilingual models based upon our results. We publish all the pre-trained and fine-tuned models freely for the research community.

Similar Papers
  • Research Article
  • Citations1

Fine-Tuning QurSim on Monolingual and Multilingual Models for Semantic Search

  • Jan 23, 2025
  • Information
  • Tania Afzal +3
  • PDF
  • Conference Article
  • Citations2

FiSSA at SemEval-2020 Task 9: Fine-tuned for Feelings

  • Jan 01, 2020
  • Bertelt Braaksma +4
  • PDF
  • Conference Article
  • Citations2

Probing Structured Pruning on Multilingual Pre-trained Models: Settings, Algorithms, and Efficiency

  • Jan 01, 2022
  • Yanyang Li +5
  • Book Chapter
  • Citations6

Training Dataset and Dictionary Sizes Matter in BERT Models: The Case of Baltic Languages

  • Jan 01, 2022
  • Matej Ulčar +1
  • Research Article
  • Citations77

Bangla-BERT: Transformer-Based Efficient Model for Transfer Learning and Language Understanding

  • Jan 01, 2022
  • IEEE Access
  • M Kowsher +5
  • PDF
  • Research Article
  • Citations18

Pre-trained transformer-based language models for Sundanese

  • Apr 13, 2022
  • Journal of Big Data
  • Wilson Wongso +2
  • Research Article
  • Citations4

Cross-lingual dependency parsing for a language with a unique script

  • Sep 09, 2024
  • Natural Language Processing
  • He Zhou +2
  • PDF
  • Conference Article
  • Citations5

Grapheme-to-Phoneme Conversion with a Multilingual Transformer Model

  • Jan 01, 2020
  • Omnia Elsaadany +1
  • PDF
  • Research Article
  • Citations10

AGI-P: A Gender Identification Framework for Authorship Analysis Using Customized Fine-Tuning of Multilingual Language Model

  • Jan 01, 2024
  • IEEE Access
  • Raheem Sarwar +6
  • Research Article

Cross-lingual Training for Multiple-Choice Question Answering

  • Oct 22, 2020
  • Procesamiento Del Lenguaje Natural
  • Guillermo Echegoyen +2
  • Conference Article

RBG-AI: Benefits of Multilingual Language Models for Low-Resource Languages

  • Jan 01, 2025
  • Barathi Ganesh Hb +1
  • PDF
  • Conference Article
  • Citations3

Evaluating Cross-Lingual Transfer Learning Approaches in Multilingual Conversational Agent Models

  • Jan 01, 2020
  • Lizhen Tan +1
  • Conference Article
  • Citations7

Can Multilingual Transformers Fight the COVID-19 Infodemic?

  • Jan 01, 2021
  • Lasitha Uyangodage +2
  • Conference Article

Beyond Multilinguality: Typological Limitations in Multilingual Models for Meitei Language

  • Jan 01, 2026
  • Badal Nyalang
  • Video Transcripts

Generalising Multilingual Concept-to-Text NLG with Language Agnostic Delexicalisation

  • Aug 01, 2021
  • Underline Science Inc.
  • Gerasimos Lampouras +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.