An Investigation into Mathematical Programming for Finite Horizon Decentralized POMDPs-Reference-Cited by-同舟云学术

An Investigation into Mathematical Programming for Finite Horizon Decentralized POMDPs

Published:2010-03-26 Issue: Volume:37 Page:329-396
ISSN:1076-9757
Container-title:Journal of Artificial Intelligence Research
language:
Short-container-title:jair

Author:

Aras R.,Dutech A.

Abstract

Decentralized planning in uncertain environments is a complex task generally dealt with by using a decision-theoretic approach, mainly through the framework of Decentralized Partially Observable Markov Decision Processes (DEC-POMDPs). Although DEC-POMDPS are a general and powerful modeling tool, solving them is a task with an overwhelming complexity that can be doubly exponential. In this paper, we study an alternate formulation of DEC-POMDPs relying on a sequence-form representation of policies. From this formulation, we show how to derive Mixed Integer Linear Programming (MILP) problems that, once solved, give exact optimal solutions to the DEC-POMDPs. We show that these MILPs can be derived either by using some combinatorial characteristics of the optimal solutions of the DEC-POMDPs or by using concepts borrowed from game theory. Through an experimental validation on classical test problems from the DEC-POMDP literature, we compare our approach to existing algorithms. Results show that mathematical programming outperforms dynamic programming but is less efficient than forward search, except for some particular problems. The main contributions of this work are the use of mathematical programming for DEC-POMDPs and a better understanding of DEC-POMDPs and of their solutions. Besides, we argue that our alternate representation of DEC-POMDPs could be helpful for designing novel algorithms looking for approximate solutions to DEC-POMDPs.

Publisher

AI Access Foundation

Subject

Artificial Intelligence

Cited by 11 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Retail Investment under Hidden Business Cycle;SSRN Electronic Journal;2024

2. Misinformation and Disinformation in Modern Warfare;Operations Research;2022-05

3. Integer Programming on the Junction Tree Polytope for Influence Diagrams;INFORMS Journal on Optimization;2020-07

4. Multi-Agent Planning under Uncertainty with Monte Carlo Q-Value Function;Applied Sciences;2019-04-04

5. Controlling a Fleet of Unmanned Aerial Vehicles to Collect Uncertain Information in a Threat Environment;Operations Research;2017-06