← 返回论文列表 📄 下载原文 PDF  ISSCC 2018 · 31.4
ISSCC 2018Session 31 · COMPUTATION IN MEMORY FOR MACHINE LEARNINGAI / ML65nm

A 65nm 1Mb Nonvolatile Computing-in-Memory ReRAM Macro with Sub-16ns Multiply-andAccumulate for Binary DNN AI Edge Processors

⚡ 本页包含 AI 生成的分析内容,仅供参考

📋 论文概要

该论文提出了一种基于65nm工艺的1Mb非易失性存内计算ReRAM宏单元,用于二进制深度神经网络(DNN)的边缘AI设备。通过亚16ns的乘累加操作,解决了传统冯·诺依曼架构中数据搬运的能耗和延迟瓶颈,实现了低功耗、高速的神经网络推理。

💡 主要创新点

核心指标
sub-16ns multiply-and-accumulate
工艺节点
65nm
重要性
发表年份
ISSCC 2018

🏷 关键词

ReRAM存内计算二进制神经网络边缘AI非易失性存储器

📄 原文摘要

Cheng-Han Yang, Cheng-Xin Xue, En-Yu Yang, Yen-Kai Chen, Yun-Sheng Chang, Tzu-Hsiang Hsu, Ya-Chin King, Chorng-Jung Lin, Ren-Shuo Liu, Chih-Cheng Hsieh, Kea-Tiong Tang, Meng-Fan Chang National Tsing Hua University, Hsinchu, Taiwan Many artificial intelligence (AI) edge devices use nonvolatile memory (NVM) to store the weights for the neural network (trained off-line on an AI server), and require lowenergy and fast I/O accesses. The deep neural networks (DNN) used by AI processors [1,2] commonly require p-layers of a convolutional neural network (CNN) and q-layers of a fully-connected network (FCN). Current DNN processors that use a conventional (von-Neumann) memory structure are limited by high access latencies, I/O energy consumption, and hardware costs. Large working data sets result in heavy accesses across the memory hierarchy, moreover large amounts of intermediate data are also generated due to the large number of multiply-and-accumulate (MAC) operations for

👥 作者与机构

Wei-Hao Chen, Kai-Xiang Li, Wei-Yu Lin, Kuo-Hsiang Hsu, Pin-Yi Li,

分类:AI / ML · 年份:ISSCC 2018