Deep neural network acceleration framework under hardware uncertainty

2018 19th International Symposium on Quality Electronic Design (ISQED) Pub Date : 2018-03-13 DOI:10.1109/ISQED.2018.8357318

M. Imani, Pushen Wang, T. Simunic

引用次数: 10

Abstract

Deep Neural Networks (DNNs) are known as effective model to perform cognitive tasks. However, DNNs are computationally expensive in both train and inference modes as they require the precision of floating point operations. Although, several prior work proposed approximate hardware to accelerate DNNs inference, they have not considered the impact of training on accuracy. In this paper, we propose a general framework called FramNN, which adjusts DNN training model to make it appropriate for underlying hardware. To accelerate training FramNN applies adaptive approximation which dynamically changes the level of hardware approximation depending on the DNN error rate. We test the efficiency of the proposed design over six popular DNN applications. Our evaluation shows that in inference, our design can achieve 1.9× energy efficiency improvement and 1.7× speedup while ensuring less than 1% quality loss. Similarly, in training mode FramNN can achieve 5.0× energy-delay product improvement as compared to baseline AMD GPU.

查看原文本刊更多论文

硬件不确定性下的深度神经网络加速框架

深度神经网络(dnn)被认为是执行认知任务的有效模型。然而，dnn在训练和推理模式下的计算成本都很高，因为它们需要浮点运算的精度。虽然之前的一些工作提出了近似硬件来加速dnn推理，但他们没有考虑训练对准确性的影响。在本文中，我们提出了一个称为FramNN的通用框架，该框架调整DNN训练模型以使其适合底层硬件。为了加速训练，采用自适应逼近，根据深度神经网络的错误率动态改变硬件逼近的水平。我们在六个流行的深度神经网络应用中测试了所提出设计的效率。我们的评估表明，在推理中，我们的设计可以实现1.9倍的能效提升和1.7倍的加速，同时保证不到1%的质量损失。同样，在训练模式下，与基准AMD GPU相比，FramNN可以实现5.0倍的能量延迟产品改进。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

2018 19th International Symposium on Quality Electronic Design (ISQED)

自引率

0.00%

发文量