Late Breaking Results: Building an On-Chip Deep Learning Memory Hierarchy Brick by Brick

2020 57th ACM/IEEE Design Automation Conference (DAC) Pub Date : 2020-07-01 DOI:10.1109/DAC18072.2020.9218728

Isak Edo Vivancos, Sayeh Sharify, M. Nikolic, Ciaran Bannon, M. Mahmoud, Alberto Delmas Lascorz, Andreas Moshovos

引用次数: 0

Abstract

Data accesses between on- and off-chip memories account for a large fraction of overall energy consumption during inference with deep learning networks. We present Boveda, a lossless on-chip memory compression technique for neural networks operating on fixed-point values. Boveda reduces the datawidth used per block of values to be only as long as necessary: since most values are of small magnitude Boveda drastically reduces their footprint. Boveda can be used to increase the effective on-chip capacity, to reduce off-chip traffic, or to reduce the on-chip memory capacity needed to achieve a performance/energy target. Boveda reduces total model footprint to 53%.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

最新突破性成果:一块砖一块砖地构建片上深度学习内存层次结构

在使用深度学习网络进行推理时，片内和片外存储器之间的数据访问占总能耗的很大一部分。我们提出了Boveda，一种用于在定点值上操作的神经网络的无损片上存储压缩技术。Boveda减少了每个值块使用的数据宽度，只要有必要:因为大多数值都是小幅度的，Boveda大大减少了它们的占用。Boveda可用于增加有效的片上容量，减少片外流量，或减少实现性能/能量目标所需的片上存储器容量。Boveda将模型的总足迹减少到53%。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

2020 57th ACM/IEEE Design Automation Conference (DAC)

自引率

0.00%

发文量