• Home
  • Search
  • Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis
  • https://doi.org/10.48550/arxiv.2409.20059Copy DOI Icon

Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis

Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

Neural metrics for machine translation (MT) evaluation have become increasingly prominent due to their superior correlation with human judgments compared to traditional lexical metrics. Researchers have therefore utilized neural metrics through quality-informed decoding strategies, achieving better results than likelihood-based methods. With the rise of Large Language Models (LLMs), preference-based alignment techniques have gained attention for their potential to enhance translation quality by optimizing model weights directly on preferences induced by quality estimators. This study focuses on Contrastive Preference Optimization (CPO) and conducts extensive experiments to evaluate the impact of preference-based alignment on translation quality. Our findings indicate that while CPO consistently outperforms Supervised Fine-Tuning (SFT) on high-quality data with regard to the alignment metric, it may lead to instability across downstream evaluation metrics, particularly between neural and lexical ones. Additionally, we demonstrate that relying solely on the base model for generating candidate translations achieves performance comparable to using multiple external systems, while ensuring better consistency across downstream metrics.

Similar Papers
  • Research Article
  • Citations8

Guiding automatic MT evaluation by means of linguistic features

  • Sep 25, 2016
  • Digital Scholarship in the Humanities
  • Elisabet Comelles +2
  • Book Chapter

A Naïve Automatic MT Evaluation Method without Reference Translations

  • Jan 01, 2011
  • Junjie Jiang +2
  • Research Article
  • Citations18

Automatic assessment of spoken-language interpreting based on machine-translation evaluation metrics

  • Mar 04, 2022
  • Interpreting
  • Xiaolei Lu +1
  • Research Article
  • Citations2

Conformalizing Machine Translation Evaluation

  • Nov 18, 2024
  • Transactions of the Association for Computational Linguistics
  • Chrysoula Zerva +1
  • Research Article

THE ANALISYS OF NEURAL MACHINE TRANSLATION SYSTEMS USING AUTOMATED MACHINE TRANSLATION EVALUATION METRICS

  • Dec 29, 2024
  • Social’no-ekonomiceskoe upravlenie: teoria i praktika
  • E S Oshanova +1
  • PDF
  • Conference Article
  • Citations12

A fuzzier approach to machine translation evaluation: A pilot study on post-editing productivity and automated metrics in commercial settings

  • Jan 01, 2015
  • Carla Parra Escartín +1
  • Research Article

How different prompts affect GPT-5s Chinese-to-English translation performance of government work reports

  • Jan 08, 2026
  • Advances in Humanities Research
  • Jingjing Feng
  • Research Article
  • Citations4

Machine-Learning-based English Quranic Translation: An Evaluation of ChatGPT

  • Aug 15, 2024
  • International Journal of Linguistics, Literature and Translation
  • Ismail Dahia +1
  • Research Article
  • Citations13

On the role of the UMLS in supporting diagnosis generation proposed by Large Language Models

  • Aug 13, 2024
  • Journal of Biomedical Informatics
  • Majid Afshar +4
  • Research Article

#2924 Comparison of large language models and traditional natural language processing techniques in predicting arteriovenous fistula failure

  • May 23, 2024
  • Nephrology Dialysis Transplantation
  • Suman Lama +6
  • Research Article

Mind the Language Gap in Digital Humanities: LLM-Aided Translation of SKOS Thesauri

  • Nov 21, 2025
  • Anthology of Computers and the Humanities
  • Felix Kraus +3
  • Book Chapter
  • Citations10

Textual Entailment Using Different Similarity Metrics

  • Jan 01, 2015
  • Tanik Saikh +3
  • PDF
  • Research Article
  • Citations2

Chinese Translation Errors in English Machine Translation Based on Wireless Sensor Network Communication Algorithm

  • Mar 20, 2022
  • Wireless Communications and Mobile Computing
  • Jinshun Long +1
  • Video Transcripts

Macro-Average: Rare Types Are Important Too

  • May 25, 2021
  • Underline Science Inc.
  • Weiqiu You +3
  • Research Article
  • Citations1

Overview of deep learning and large language models in machine translation: a special perspective on the Arabic language

  • Jun 23, 2025
  • Journal of Electrical Systems and Information Technology
  • Sanaa Abou Elhamayed +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.