Author:
Ahn Hyun-Soo,Righter Rhonda
Abstract
We give a very general reformulation of multi-actor Markov decision processes and show that there is a tendency for the actors to take the same action whenever possible. This considerably reduces the complexity of the problem, either facilitating numerical computation of the optimal policy or providing a basis for a heuristic.
Publisher
Cambridge University Press (CUP)
Subject
Statistics, Probability and Uncertainty,General Mathematics,Statistics and Probability
Reference13 articles.
1. [8] Koole K. and Righter R. (2004). Resource allocation in grid computing. Work in progress.
2. Server Assignment Policies for Maximizing the Steady-State Throughput of Finite Queueing Systems
3. Dynamic Scheduling of a Multiclass Queue: Discount Optimality
4. [5] Kaufman D. , Ahn H.-S. and Lewis M. E. (2004). On the introduction of agile, temporary workers into a tandem queueing system. Work in progress.
Cited by
6 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献