Reinforcement Learning Based Cooperative Coded Caching Under Dynamic Popularities in Ultra-Dense Networks

被引：26

作者：

Gao, Shen ^{[1
,2
]}

Dong, Peihao ^{[1
]}

Pan, Zhiwen ^{[1
,2
]}

Li, Geoffrey Ye ^{[3
]}

机构：

[1] Southeast Univ, Natl Mobile Commun Res Lab, Nanjing 210096, Peoples R China

[2] Purple Mt Labs, Nanjing 211100, Peoples R China

[3] Georgia Inst Technol, Sch Elect & Comp Engn, Atlanta, GA 30332 USA

来源：

IEEE TRANSACTIONS ON VEHICULAR TECHNOLOGY | 2020年 / 69卷 / 05期

关键词：

Ultra-dense network; reinforcement learning; cooperative coded caching; popularity dynamics; WIRELESS; DELIVERY; DESIGN; TRANSMISSION; MIMO;

D O I：

10.1109/TVT.2020.2979918

中图分类号：

TM [电工技术]; TN [电子技术、通信技术];

学科分类号：

0808 ; 0809 ;

摘要：

For ultra-dense networks with wireless backhaul, caching strategy at small base stations (SBSs), usually with limited storage, is critical to meet massive high data rate requests. Since the content popularity profile varies with time in an unknown way, we exploit reinforcement learning (RL) to design a cooperative caching strategy with maximum-distance separable (MDS) coding. We model the MDS coding based cooperative caching as a Markov decision process to capture the popularity dynamics and maximize the long-term expected cumulative traffic load served directly by the SBSs without accessing the macro base station. For the formulated problem, we first find the optimal solution for a small-scale system by embedding the cooperative MDS coding into Q-learning. To cope with the large-scale case, we approximate the state-action value function heuristically. The approximated function includes only a small number of learnable parameters and enables us to propose a fast and efficient action-selection approach, which dramatically reduces the complexity. Numerical results verify the optimality/near-optimality of the proposed RL based algorithms and show the superiority compared with the baseline schemes. They also exhibit good robustness to different environments.

引用

页码：5442 / 5456

页数：15

共 47 条

[1] What Will 5G Be? [J].

Andrews, Jeffrey G. ;

Buzzi, Stefano ;

Choi, Wan ;

Hanly, Stephen V. ;

Lozano, Angel ;

Soong, Anthony C. K. ;

Zhang, Jianzhong Charlie .

IEEE JOURNAL ON SELECTED AREAS IN COMMUNICATIONS, 2014, 32 (06) :1065-1082

[2]

[Anonymous], 2011, CISC VIS NETW IND GL

[3]

[Anonymous], 2016, GOVT RESOLUTION REPU

[4]

[Anonymous], 2014, M2320 ITUR

[5]

[Anonymous], 2015, IEEE T SIGNAL PROCES, DOI DOI 10.1109/TSP.2014.2367473

[6]

Bastug E, 2016, IEEE INT SYMP INFO, P285, DOI 10.1109/ISIT.2016.7541306

[7]

Breslau L, 1999, IEEE INFOCOM SER, P126, DOI 10.1109/INFCOM.1999.749260

[8] Cooperative Caching and Transmission Design in Cluster-Centric Small Cell Networks [J].

Chen, Zheng ;

Lee, Jemin ;

Quek, Tony Q. S. ;

Kountouris, Marios .

IEEE TRANSACTIONS ON WIRELESS COMMUNICATIONS, 2017, 16 (05) :3401-3415

[9] Analysis and Optimization of Caching and Multicasting in Large-Scale Cache-Enabled Heterogeneous Wireless Networks [J].

Cui, Ying ;

Jiang, Dongdong .

IEEE TRANSACTIONS ON WIRELESS COMMUNICATIONS, 2017, 16 (01) :250-264

[10] Wireless Backhaul Networks: Capacity Bound, Scalability Analysis and Design Guidelines [J].

Dhillon, Harpreet S. ;

Caire, Giuseppe .

IEEE TRANSACTIONS ON WIRELESS COMMUNICATIONS, 2015, 14 (11) :6043-6056

← 1 2 3 4 5 →