Incentive-Aware Recommender Systems in Two-Sided Markets-Reference-Cited by-同舟云学术

Incentive-Aware Recommender Systems in Two-Sided Markets

Published:2024-07-31 Issue:4 Volume:2 Page:1-38
ISSN:2770-6699
Container-title:ACM Transactions on Recommender Systems
language:en
Short-container-title:ACM Trans. Recomm. Syst.

Author:

Dai Xiaowu¹^ORCID,Xu Wenlu¹^ORCID,Qi Yuan²^ORCID,Jordan Michael³^ORCID

Affiliation:

1. University of California, Los Angeles, Los Angeles, United States

2. Fudan University, Shanghai, China and Ant Group, Hangzhou, China

3. UC Berkeley, Berkeley, United States and Ant Group, Hangzhou, China

Abstract

Online platforms in the Internet Economy commonly incorporate recommender systems that recommend products (or “arms”) to users (or “agents”). A key challenge in this domain arises from myopic agents who are naturally incentivized to exploit by choosing the optimal arm based on current information, rather than exploring various alternatives to gather information that benefits the collective. We propose a new recommender system that aligns with agents’ incentives while achieving asymptotically optimal performance, as measured by regret in repeated interactions. Our framework models this incentive-aware system as a multi-agent bandit problem in two-sided markets, where the interactions of agents and arms are facilitated by recommender systems on online platforms. This model incorporates incentive constraints induced by agents’ opportunity costs. In scenarios where opportunity costs are known to the platform, we show the existence of an incentive-compatible recommendation algorithm. This algorithm pools recommendations between a genuinely good arm and an unknown arm using a randomized and adaptive strategy. Moreover, when these opportunity costs are unknown, we introduce an algorithm that randomly pools recommendations across all arms, utilizing the cumulative loss from each arm as feedback for strategic exploration. We demonstrate that both algorithms satisfy an ex-post fairness criterion, which protects agents from over-exploitation. All code for using the proposed algorithms and reproducing results is made available on GitHub.

Publisher

Association for Computing Machinery (ACM)

Link

https://dl.acm.org/doi/pdf/10.1145/3674158

Reference50 articles.

1. Learning From Reviews: The Selection Effect and the Speed of Learning

2. Robert J. Aumann, Michael Maschler, and Richard E. Stearns. 1995. Repeated Games with Incomplete Information. MIT Press.

3. Economic recommendation systems;Bahar Gal;Proceedings of the 16th ACM Conference on Electronic Commerce (EC),2015

4. A Simple Model of Herd Behavior

5. Mostly Exploration-Free Algorithms for Contextual Bandits