• Home
  • Search
  • Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation
  • Cite Icon5
  • https://doi.org/10.1145/3640457.3688142Copy DOI Icon

Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation

  • Oct 8, 2024
  • David Austin +3 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Designing preference elicitation (PE) methodologies that can quickly ascertain a user’s top item preferences in a cold-start setting is a key challenge for building effective and personalized conversational recommendation (ConvRec) systems. While large language models (LLMs) enable fully natural language (NL) PE dialogues, we hypothesize that monolithic LLM NL-PE approaches lack the multi-turn, decision-theoretic reasoning required to effectively balance the exploration and exploitation of user preferences towards an arbitrary item set. In contrast, traditional Bayesian optimization PE methods define theoretically optimal PE strategies, but cannot generate arbitrary NL queries or reason over content in NL item descriptions – requiring users to express preferences via ratings or comparisons of unfamiliar items. To overcome the limitations of both approaches, we formulate NL-PE in a Bayesian Optimization (BO) framework that seeks to actively elicit NL feedback to identify the best recommendation. Key challenges in generalizing BO to deal with natural language feedback include determining: (a) how to leverage LLMs to model the likelihood of NL preference feedback as a function of item utilities, and (b) how to design an acquisition function for NL BO that can elicit preferences in the infinite space of language. We demonstrate our framework in a novel NL-PE algorithm, PEBOL, which uses: 1) Natural Language Inference (NLI) between user preference utterances and NL item descriptions to maintain Bayesian preference beliefs, and 2) BO strategies such as Thompson Sampling (TS) and Upper Confidence Bound (UCB) to guide LLM query generation. We numerically evaluate our methods in controlled simulations, finding that after 10 turns of dialogue, PEBOL can achieve an MRR@10 of up to 0.27 compared to the best monolithic LLM baseline’s MRR@10 of 0.17, despite relying on earlier and smaller LLMs.1

Similar Papers
  • Research Article

Evaluating gpt-4 for zero-shot classification of bleeding and clotting events: Can large language models serve as second reviewers?

  • Nov 03, 2025
  • Blood
  • Samantha Rizzo +5
  • Research Article
  • Citations51

Large language models for biomedicine: foundations, opportunities, challenges, and best practices.

  • Apr 24, 2024
  • Journal of the American Medical Informatics Association : JAMIA
  • Satya S Sahoo +8
  • Preprint Article

Simulated Selfhood in LLMs: A Behavioral Analysis of Introspective Coherence

  • Jul 26, 2025
  • Jose Augusto De Lima Prestes
  • Research Article
  • Citations1

Logical and Physical Optimizations for SQL Query Execution over Large Language Models

  • Jun 17, 2025
  • Proceedings of the ACM on Management of Data
  • Dario Satriani +5
  • Conference Article
  • Citations7

Enhanced Recommendation Combining Collaborative Filtering and Large Language Models

  • Jan 17, 2025
  • Xueting Lin +4
  • Research Article

ModuLoRA: Finetuning 2-Bit LLMs on Consumer GPUs by Integrating with Modular Quantizers.

  • Feb 01, 2024
  • Transactions on machine learning research
  • Junjie Yin +4
  • Research Article
  • Citations67

The life cycle of large language models in education: A framework for understanding sources of bias

  • Jul 12, 2024
  • British Journal of Educational Technology
  • Jinsook Lee +4
  • Research Article

#2924 Comparison of large language models and traditional natural language processing techniques in predicting arteriovenous fistula failure

  • May 23, 2024
  • Nephrology Dialysis Transplantation
  • Suman Lama +6
  • Research Article

Can LLMs effectively provide game-theoretic-based scenarios for cybersecurity?

  • Dec 11, 2025
  • Frontiers in Computer Science
  • Daniele Proverbio +5
  • Research Article
  • Citations1

Cross Category Recommendations Using LLMs

  • Mar 30, 2024
  • Darpan International Research Analysis
  • Murali Mohana Krishna Dandu +4
  • Research Article
  • Citations20

DialogueLLM: Context and emotion knowledge-tuned large language models for emotion recognition in conversations.

  • Dec 01, 2025
  • Neural networks : the official journal of the International Neural Network Society
  • Yazhou Zhang +6
  • Research Article
  • Citations2

TurkMedNLI: a Turkish medical natural language inference dataset through large language model based translation.

  • Jan 30, 2025
  • PeerJ. Computer science
  • İskender Ülgen Oğul +2
  • Conference Article

IML at SemEval-2024 Task 2: Safe Biomedical Natural Language Interference for Clinical Trials with LLM Based Ensemble Inferencing

  • Jan 01, 2024
  • Abbas Akkasi +4
  • Research Article
  • Citations25

LLM-Informed Multi-Armed Bandit Strategies for Non-Stationary Environments

  • Jun 25, 2023
  • Electronics
  • J De Curtò +5
  • PDF
  • Conference Article
  • Citations13

Unleashing the Retrieval Potential of Large Language Models in Conversational Recommender Systems

  • Oct 08, 2024
  • Ting Yang +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.