• Home
  • Search
  • Enabling On-Device Large Language Model Personalization with Self-Supervised Data Selection and Synthesis
  • Open Access IconOpen Access
  • Cite Icon13
  • https://doi.org/10.1145/3649329.3655665Copy DOI Icon

Enabling On-Device Large Language Model Personalization with Self-Supervised Data Selection and Synthesis

  • Jun 23, 2024
  • Ruiyang Qin +7 more
Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

After a large language model (LLM) is deployed on edge devices, it is desirable for these devices to learn from user-generated conversation data to generate user-specific and personalized responses in real-time. However, user-generated data usually contains sensitive and private information, and uploading such data to the cloud for annotation is not preferred if not prohibited. While it is possible to obtain annotation locally by directly asking users to provide preferred responses, such annotations have to be sparse to not affect user experience. In addition, the storage of edge devices is usually too limited to enable large-scale fine-tuning with full user-generated data. It remains an open question how to enable on-device LLM personalization, considering sparse annotation and limited on-device storage. In this paper, we propose a novel framework to select and store the most representative data online in a self-supervised way. Such data has a small memory footprint and allows infrequent requests of user annotations for further fine-tuning. To enhance fine-tuning quality, multiple semantically similar pairs of question texts and expected responses are generated using the LLM. Our experiments show that the proposed framework achieves the best user-specific content-generating capability (accuracy) and fine-tuning speed (performance) compared with vanilla baselines. To the best of our knowledge, this is the very first on-device LLM personalization framework.

Similar Papers
  • Research Article
  • Citations21

Two-Layer Retrieval-Augmented Generation Framework for Low-Resource Medical Question Answering Using Reddit Data: Proof-of-Concept Study.

  • Jan 06, 2025
  • Journal of medical Internet research
  • Sudeshna Das +21
  • PDF
  • Research Article
  • Citations3

Exploring the potential of lightweight large language models for AI-based mental health counselling task: a novel comparative study

  • Jul 02, 2025
  • Scientific Reports
  • Ritesh Maurya +4
  • Research Article

English

  • Feb 28, 2025
  • International Journal of Computer Trends and Technology
  • Dhivya Nagasubramanian
  • Conference Article
  • Citations9

Robust Implementation of Retrieval-Augmented Generation on Edge-based Computing-in-Memory Architectures

  • Oct 27, 2024
  • Ruiyang Qin +11
  • Research Article

Securing social network user data in large language model deployments: challenges and best practices

  • Sep 11, 2025
  • Cluster Computing
  • Nasir Ahmad Jalali +1
  • Research Article
  • Citations12

ROFED-LLM: Robust Federated Learning for Large Language Models in Adversarial Wireless Environments

  • Jan 01, 2026
  • IEEE Transactions on Network Science and Engineering
  • Haoyu Wang +6
  • Conference Article

TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent

  • Jan 01, 2025
  • Dominik Meier +4
  • Conference Article
  • Citations1

A First Look at LLM-powered Smartphones

  • Oct 27, 2024
  • Liangxuan Wu +4
  • Research Article
  • Citations5

Reflective Dialogues with a Humanoid Robot Integrated with an LLM and a Curated NLU System for Positive Behavioral Change in Older Adults

  • Nov 07, 2024
  • Electronics
  • Ryan Browne +17
  • Research Article

LLMGuard : Safeguarding Real-Time Inference for Large Language Models on Edge Devices

  • Feb 25, 2026
  • ACM Transactions on Software Engineering and Methodology
  • Yu Sun +4
  • Conference Article

FBQuant: FeedBack Quantization for Large Language Models

  • Sep 01, 2025
  • Yijiang Liu +6
  • Research Article
  • Citations5

GenAI at the Edge: Comprehensive Survey on Empowering Edge Devices

  • May 28, 2025
  • Proceedings of the AAAI Symposium Series
  • Mozhgan Navardi +6
  • Research Article

Refinetuning Decentralized Large Language Model for Privacy-Sensitive University Data

  • Mar 17, 2025
  • Journal of Robotics and Automation Research
  • Kilian Lorenz +6
  • Conference Article
  • Citations4

Can LLMs get help from other LLMs without revealing private information?

  • Jan 01, 2024
  • Florian Hartmann +4
  • Preprint Article

Information-Theoretic Privacy Control for Sequential Multi-Agent LLM Systems

  • Feb 13, 2026
  • arXiv (Cornell University)
  • Sadia Asif +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.