Reinforcement Learning-Based Multihop Relaying: A Decentralized Q-Learning Approach-Reference-Cited by-同舟云学术

Reinforcement Learning-Based Multihop Relaying: A Decentralized Q-Learning Approach

Published:2021-10-06 Issue:10 Volume:23 Page:1310
ISSN:1099-4300
Container-title:Entropy
language:en
Short-container-title:Entropy

Author:

Wang Xiaowei,Wang Xin

Abstract

Conventional optimization-based relay selection for multihop networks cannot resolve the conflict between performance and cost. The optimal selection policy is centralized and requires local channel state information (CSI) of all hops, leading to high computational complexity and signaling overhead. Other optimization-based decentralized policies cause non-negligible performance loss. In this paper, we exploit the benefits of reinforcement learning in relay selection for multihop clustered networks and aim to achieve high performance with limited costs. Multihop relay selection problem is modeled as Markov decision process (MDP) and solved by a decentralized Q-learning scheme with rectified update function. Simulation results show that this scheme achieves near-optimal average end-to-end (E2E) rate. Cost analysis reveals that it also reduces computation complexity and signaling overhead compared with the optimal scheme.

Funder

National Natural Science Foundation of China

Shanghai Municipal Education Commission

Publisher

MDPI AG

Subject

General Physics and Astronomy

Link

https://www.mdpi.com/1099-4300/23/10/1310/pdf

Reference19 articles.

1. Performance Analysis of Cluster-Based Multi-Hop Underlay CRNs Using Max-Link-Selection Protocol

2. Decentralized Relay Selection in Multi-User Multihop Decode-and-Forward Relay Networks