HEIF: Highly Efficient Stochastic Computing-Based Inference Framework for Deep Neural Networks

被引:54
作者
Li, Zhe [1 ]
Li, Ji [2 ]
Ren, Ao [1 ]
Cai, Ruizhe [1 ]
Ding, Caiwen [1 ]
Qian, Xuehai [2 ]
Draper, Jeffrey [2 ]
Yuan, Bo [3 ]
Tang, Jian [1 ]
Qiu, Qinru [1 ]
Wang, Yanzhi [4 ]
机构
[1] Syracuse Univ, Dept Elect Engn & Comp Sci, Syracuse, NY 13244 USA
[2] Univ Southern Calif, Dept Elect Engn, Los Angeles, CA 90089 USA
[3] CUNY City Coll, Dept Elect Engn, New York, NY 10031 USA
[4] Northeastern Univ, Dept Elect & Comp Engn, Boston, MA 02115 USA
关键词
ASIC; convolutional neural network; deep learning; energy-efficient; optimization; stochastic computing (SC);
D O I
10.1109/TCAD.2018.2852752
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
Deep convolutional neural networks (DCNNs) are one of the most promising deep learning techniques and have been recognized as the dominant approach for almost all recognition and detection tasks. The computation of DCNNs is memory intensive due to large feature maps and neuron connections, and the performance highly depends on the capability of hardware resources. With the recent trend of wearable devices and Internet of Things, it becomes desirable to integrate the DCNNs onto embedded and portable devices that require low power and energy consumptions and small hardware footprints. Recently stochastic computing (SC)-DCNN demonstrated that SC as a low-cost substitute to binary-based computing radically simplifies the hardware implementation of arithmetic units and has the potential to satisfy the stringent power requirements in embedded devices. In SC, many arithmetic operations that are resource-consuming in binary designs can be implemented with very simple hardware logic, alleviating the extensive computational complexity. It offers a colossal design space for integration and optimization due to its reduced area and soft error resiliency. In this paper, we present HEIF, a highly efficient SC-based inference framework of the large-scale DCNNs, with broad applications including (but not limited to) LeNet-5 and AlexNet, that achieves high energy efficiency and low area/ hardware cost. Compared to SC-DCNN, HEIF features: 1) the first (to the best of our knowledge) SC-based rectified linear unit activation function to catch up with the recent advances in software models and mitigate degradation in application-level accuracy; 2) the redesigned approximate parallel counter and optimized stochastic multiplication using transmission gates and inverse mirror adders; and 3) the new optimization of weight storage using clustering. Most importantly, to achieve maximum energy efficiency while maintaining acceptable accuracy, HEIF considers holistic optimizations on cascade connection of function blocks in DCNN, pipelining technique, and bit-stream length reduction. Experimental results show that in large-scale applications HEIF outperforms previous SC-DCNN by the throughput of 4.1x, by area efficiency of up to 6.5x, and achieves up to 5.6x energy improvement.
引用
收藏
页码:1543 / 1556
页数:14
相关论文
共 50 条
  • [41] PID Controller-Based Stochastic Optimization Acceleration for Deep Neural Networks
    Wang, Haoqian
    Luo, Yi
    An, Wangpeng
    Sun, Qingyun
    Xu, Jun
    Zhang, Lei
    IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2020, 31 (12) : 5079 - 5091
  • [42] Stochastic Weight Matrix-Based Regularization Methods for Deep Neural Networks
    Reizinger, Patrik
    Gyires-Toth, Balint
    MACHINE LEARNING, OPTIMIZATION, AND DATA SCIENCE, 2019, 11943 : 45 - 57
  • [43] Energy Efficient Learning With Low Resolution Stochastic Domain Wall Synapse for Deep Neural Networks
    Al Misba, Walid
    Lozano, Mark
    Querlioz, Damien
    Atulasimha, Jayasimha
    IEEE ACCESS, 2022, 10 : 84946 - 84959
  • [44] A sparsity-based stochastic pooling mechanism for deep convolutional neural networks
    Song, Zhenhua
    Liu, Yan
    Song, Rong
    Chen, Zhenguang
    Yang, Jianyong
    Zhang, Chao
    Jiang, Qing
    NEURAL NETWORKS, 2018, 105 : 340 - 345
  • [45] A soft computing-based approach for integrated training and rule extraction from artificial neural networks: DIFACONN-miner
    Oezbakir, Lale
    Baykasoglu, Adil
    Kulluk, Sinem
    APPLIED SOFT COMPUTING, 2010, 10 (01) : 304 - 317
  • [46] Edge computing-based internet of medical things for healthcare using deep learning
    Sathyaveti, Himabindu
    Gomathy, C.
    INTERNATIONAL JOURNAL OF EMBEDDED SYSTEMS, 2023, 16 (02) : 117 - 125
  • [47] Scalable Stochastic-Computing Accelerator for Convolutional Neural Networks
    Sim, Hyeonuk
    Dong Nguyen
    Lee, Jongeun
    Choi, Kiyoung
    2017 22ND ASIA AND SOUTH PACIFIC DESIGN AUTOMATION CONFERENCE (ASP-DAC), 2017, : 696 - 701
  • [48] A Survey of Stochastic Computing Neural Networks for Machine Learning Applications
    Liu, Yidong
    Liu, Siting
    Wang, Yanzhi
    Lombardi, Fabrizio
    Han, Jie
    IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2021, 32 (07) : 2809 - 2824
  • [49] Reliable and Efficient Multimedia Service Optimization for Edge Computing-Based 5G Networks: Game Theoretic Approaches
    Cao, Tengfei
    Xu, Changqiao
    Du, Junping
    Li, Yawen
    Xiao, Han
    Gong, Changhui
    Zhong, Lujie
    Niyato, Dusit
    IEEE TRANSACTIONS ON NETWORK AND SERVICE MANAGEMENT, 2020, 17 (03): : 1610 - 1625
  • [50] Accelerating Inference of Convolutional Neural Networks Using In-memory Computing
    Dazzi, Martino
    Sebastian, Abu
    Benini, Luca
    Eleftheriou, Evangelos
    FRONTIERS IN COMPUTATIONAL NEUROSCIENCE, 2021, 15