Static Scheduling of Weight Programming for DNN Acceleration with Resource Constrained PIM
Xin Gao, Hongyue Wang, Yiyan Chen, Yuhao Zhang, Zhaoyan Shen, Lei Ju
Shandong University Peng Cheng Laboratory
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Most existing architectural studies on ReRAM-based processing-in-memory (PIM) DNN accelerators assume that all weights of the DNN can be mapped to the crossbar at once. However, these studies are over-idealized. ReRAM crossbar resources for calculation are limited because of technological limitations, so multiple weight mapping procedures are required during the inference process. In this article, we propose a static scheduling framework which generates the mapping between DNN weights and ReRAM cells with minimum runtime weight programming cost. We first build a ReRAM crossbar programming latency model by simultaneously considering the DNN weight patterns, ReRAM programming operations, and PIM architecture characteristics. Then, the model is used in the searching process to obtain an optimized weight-to-OU mapping table with minimum online programming latency. Finally, an OU scheduler is used to coordinate the activation sequences of OUs in the crossbars to perform the inference computation correctly. Evaluation results show the proposed framework significantly reduces the weight programming overhead and the overall inference latency for various DNN models with different input datasets.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
工程Advanced Memory and Neural Computing
Ferroelectric and Negative Capacitance Devices · Parallel Computing and Optimization Techniques
参考文献 46
此处列出前 3 条
引用本文 3
按被引量排序,此处列出前 3 条