TensorX
返回文献探索

Paper · arXiv 2305.05973

Privacy-Preserving Recommender Systems with Synthetic Query Generation using Differentially Private Large Language Models

Aldo Gael Carranza, Rezsa Farahani, Natalia Ponomareva, Alex Kurakin, Matthew Jagielski, Milad Nasr

1 upvotesMay 10, 2023arXiv 预印本
AI 摘要

A method using differentially private large language models for generating private synthetic queries enhances the security and effectiveness of deep retrieval models in recommender systems.

differentially private (DP)large language models (LLMs)DP trainingfine-tuningsynthetic queriesdeep retrieval modelsquery-level privacy guarantees

Abstract

We propose a novel approach for developing privacy-preserving large-scale recommender systems using differentially private (DP) large language models (LLMs) which overcomes certain challenges and limitations in DP training these complex systems. Our method is particularly well suited for the emerging area of LLM-based recommender systems, but can be readily employed for any recommender systems that process representations of natural language inputs. Our approach involves using DP training methods to fine-tune a publicly pre-trained LLM on a query generation task. The resulting model can generate private synthetic queries representative of the original queries which can be freely shared for any downstream non-private recommendation training procedures without incurring any additional privacy cost. We evaluate our method on its ability to securely train effective deep retrieval models, and we observe significant improvements in their retrieval quality without compromising query-level privacy guarantees compared to methods where the retrieval models are directly DP trained.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Privacy-Preserving Recommender Systems with Synthetic Query Generation using Differentially Private Large Language Models | TensorX