İçeriğe geç
akaturk Akademik ölçüm

Makale detayı · 2022

A Heterogeneous and Programmable Compute-In-Memory Accelerator Architecture for Analog-AI Using Dense 2-D Mesh

Dergi

IEEE Transactions on Very Large Scale Integration (VLSI) Systems
OpenAlex SJR Q1 JCR Q2 Atıf 74 Üst %10 Yüzdelik 96.4% FWCI 4.86
Yıl
2022
Tür
article

Veri kaynağı ayrımı

  • YÖKSİS dergi adı IEEE Transactions on Very Large Scale Integration (VLSI) Systems
  • OpenAlex OpenAlex zenginleştirmesi (özet, atıf, konular)

Özet

OpenAlex · İngilizce

We introduce a highly heterogeneous and programmable compute-in-memory (CIM) accelerator architecture for deep neural network (DNN) inference. This architecture combines spatially distributed CIM memory array “tiles” for weight-stationary, energy-efficient multiply–accumulate (MAC) operations, together with heterogeneous special-function compute cores for auxiliary digital computation. Massively parallel vectors of neuron activation data are exchanged over short distances using a dense and efficient circuit-switched 2-D mesh, offering full end-to-end support for a wide range of DNN workloads, including CNNs, long-short-term-memory (LSTM), and transformers. We discuss the design of the “analog fabric”—the 2-D grid of tiles and compute cores interconnected by the 2-D mesh—and address the efficiency in both mapping of DNNs onto the hardware and in pipelining of various DNN workloads across a range of batch sizes. We show, for the first time, system-level assessments using projected component parameters for a realistic “analog AI” system, based on dense crossbar arrays of low-power nonvolatile analog memory elements, while incorporating a single common analog fabric design that can scale to large networks by introducing data transport between multiple analog AI chips. Our performance estimates for several networks, including large LSTM and bidirectional encoder representations from transformers (BERT), show highly competitive throughput while offering$40\times $–$140\times $higher energy efficiency than NVIDIA A100—thus illustrating the strong promise of analog AI and the proposed architecture for DNN inference applications.

Konular

Atıflar

OpenAlex cited_by_count. WoS veya Scopus atıf sayısı değildir; o kaynaklar için ayrı kolon yoktur.

74 atıf

OpenAlex cited_by_count (önbellek / veritabanı)

Yazarlar

Yazar bilgisi yok.