Budget-Constrained Multi-Armed Bandits With Multiple Plays-Reference-Cited by-同舟云学术

Budget-Constrained Multi-Armed Bandits With Multiple Plays

Published:2018-04-29 Issue:1 Volume:32 Page:
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Zhou Datong,Tomlin Claire

Abstract

We study the multi-armed bandit problem with multiple plays and a budget constraint for both the stochastic and the adversarial setting. At each round, exactly K out of N possible arms have to be played (with 1 ≤ K <= N). In addition to observing the individual rewards for each arm played, the player also learns a vector of costs which has to be covered with an a-priori defined budget B. The game ends when the sum of current costs associated with the played arms exceeds the remaining budget. Firstly, we analyze this setting for the stochastic case, for which we assume each arm to have an underlying cost and reward distribution with support [cmin, 1] and [0, 1], respectively. We derive an Upper Confidence Bound (UCB) algorithm which achieves O(NK4 log B) regret. Secondly, for the adversarial case in which the entire sequence of rewards and costs is fixed in advance, we derive an upper bound on the regret of order O(√NB log(N/K)) utilizing an extension of the well-known Exp3 algorithm. We also provide upper bounds that hold with high probability and a lower bound of order Ω((1 – K/N) √NB/K).

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 7 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Client selection for federated learning using combinatorial multi-armed bandit under long-term energy constraint;Computer Networks;2024-08

2. Combinatorial Incentive Mechanism for Bundling Spatial Crowdsourcing with Unknown Utilities;IEEE INFOCOM 2024 - IEEE Conference on Computer Communications;2024-05-20

3. Incentive-driven long-term optimization for hierarchical federated learning;Computer Networks;2023-10

4. You Can Trade Your Experience in Distributed Multi-Agent Multi-Armed Bandits;2023 IEEE/ACM 31st International Symposium on Quality of Service (IWQoS);2023-06-19

5. Multi-armed Bandit with Time-variant Budgets;2023 4th International Conference on Computer Engineering and Application (ICCEA);2023-04-07