On the use of Deep Autoencoders for Efficient Embedded Reinforcement Learning

被引:7
作者
Prakash, Bharat [1 ]
Horton, Mark [1 ]
Waytowich, Nicholas R. [2 ]
Hairston, William David [2 ]
Oates, Tim [1 ]
Mohsenin, Tinoosh [1 ]
机构
[1] Univ Maryland Baltimore Cty, Baltimore, MD 21228 USA
[2] US Army, Res Lab, Washington, DC 20310 USA
来源
GLSVLSI '19 - PROCEEDINGS OF THE 2019 ON GREAT LAKES SYMPOSIUM ON VLSI | 2019年
关键词
Autoencoders; Neural networks; Reinforcement learning; Embedded devices; Deep learning;
D O I
10.1145/3299874.3319493
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
In autonomous embedded systems, it is often vital to reduce the amount of actions taken in the real world and energy required to learn a policy. Training reinforcement learning agents from high dimensional image representations can be very expensive and time consuming. Autoencoders are deep neural network used to compress high dimensional data such as pixelated images into small latent representations. This compression model is vital to efficiently learn policies, especially when learning on embedded systems. We have implemented this model on the NVIDIA Jetson TX2 embedded GPU, and evaluated the power consumption, throughput, and energy consumption of the autoencoders for various CPU/GPU core combinations, frequencies, and model parameters. Additionally, we have shown the reconstructions generated by the autoencoder to analyze the quality of the generated compressed representation and also the performance of the reinforcement learning agent. Finally, we have presented an assessment of the viability of training these models on embedded systems and their usefulness in developing autonomous policies. Using autoencoders, we were able to achieve 4-5 x improved performance compared to a baseline RL agent with a convolutional feature extractor, while using less than 2W of power.
引用
收藏
页码:507 / 512
页数:6
相关论文
共 12 条
  • [11] Warnell G, 2018, AAAI CONF ARTIF INTE, P1545
  • [12] Waytowich Nicholas R., 2018, ABS180809572V CORR