Is ChatGPT Fair for Recommendation? Evaluating Fairness in Large Language Model Recommendation

The remarkable achievements of Large Language Models (LLMs) have led to the emergence of a novel recommendation paradigm -- Recommendation via LLM (RecLLM). Nevertheless, it is important to note that LLMs may contain social prejudices, and therefore, the fairness of recommendations made by RecLLM requires further investigation. To avoid the potential risks of RecLLM, it is imperative to evaluate the fairness of RecLLM with respect to various sensitive attributes on the user side. Due to the differences between the RecLLM paradigm and the traditional recommendation paradigm, it is problematic to directly use the fairness benchmark of traditional recommendation. To address the dilemma, we propose a novel benchmark called Fairness of Recommendation via LLM (FaiRLLM). This benchmark comprises carefully crafted metrics and a dataset that accounts for eight sensitive attributes1 in two recommendation scenarios: music and movies. By utilizing our FaiRLLM benchmark, we conducted an evaluation of ChatGPT and discovered that it still exhibits unfairness to some sensitive attributes when generating recommendations. Our code and dataset can be found at https://github.com/jizhi-zhang/FaiRLLM.

Data Mining: Conceptsand TechniquesData Mining: Concepts and TechniquesThe Unfairness ofPopularity Bias in…The Unfairness of Popularity Bias in RecommendationToward Pareto EfficientFairness-Utility…Toward Pareto Efficient Fairness-Utility Trade-off in Recommendation through Reinforcement LearningOPT: Open Pre-trainedTransformer Language…OPT: Open Pre-trained Transformer Language ModelsTraining language modelsto follow instructions…Training language models to follow instructions with human feedbackA Survey on the Fairnessof Recommender SystemsA Survey on the Fairness of Recommender SystemsChatGPT: Fundamentals,Applications and Social…ChatGPT: Fundamentals, Applications and Social ImpactsGenerativeRecommendation: Towards…Generative Recommendation: Towards Next-generation Recommender ParadigmLLaMA: Open andEfficient Foundation…LLaMA: Open and Efficient Foundation Language ModelsEvaluating theEffectiveness of Large…Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations (Short Paper)Exploring the UpperLimits of Text-Based…Exploring the Upper Limits of Text-Based Collaborative Filtering Using Large Language Models: Discoveries and InsightsA Survey of LargeLanguage ModelsA Survey of Large Language ModelsTALLRec: An Effectiveand Efficient Tuning…TALLRec: An Effective and Efficient Tuning Framework to Align Large Language Model with RecommendationA Large Language ModelEnhanced Conversational…A Large Language Model Enhanced Conversational Recommender SystemLarge Language Modelsare Zero-Shot Rankers…Large Language Models are Zero-Shot Rankers for Recommender SystemsHow Can RecommenderSystems Benefit from…How Can Recommender Systems Benefit from Large Language Models: A SurveyProspect PersonalizedRecommendation on Large…Prospect Personalized Recommendation on Large Language Model-based Agent PlatformReLLa:Retrieval-enhanced Larg…ReLLa: Retrieval-enhanced Large Language Models for Lifelong Sequential Behavior Comprehension in RecommendationRecommender Systems inthe Era of Large…Recommender Systems in the Era of Large Language Models (LLMs)Lifelong PersonalizedLow-Rank Adaptation of…Lifelong Personalized Low-Rank Adaptation of Large Language Models for RecommendationA Survey of GenerativeSearch and…A Survey of Generative Search and Recommendation in the Era of Large Language ModelsUnderstanding Biases inChatGPT-based…Understanding Biases in ChatGPT-based Recommender Systems: Provider Fairness, Temporal Stability, and RecencyA Survey on Evaluationof Large Language ModelsA Survey on Evaluation of Large Language ModelsA Bi-Step GroundingParadigm for Large…A Bi-Step Grounding Paradigm for Large Language Models in Recommendation SystemsIs ChatGPT Fair forRecommendation?…Is ChatGPT Fair for Recommendation? Evaluating Fairness in Large Language Model RecommendationEarlier referencesFocus paperCiting papersOlderNewer

Click a node to pin it, click the empty canvas to go back to this paper, or hover to preview. Open a node’s page from its title.