Makale detayı · 2022
A Heterogeneous and Programmable Compute-In-Memory Accelerator Architecture for Analog-AI Using Dense 2-D Mesh
Dergi
IEEE Transactions on Very Large Scale Integration (VLSI) Systems- Yıl
- 2022
- Tür
- article
Veri kaynağı ayrımı
- YÖKSİS dergi adı IEEE Transactions on Very Large Scale Integration (VLSI) Systems
- OpenAlex OpenAlex zenginleştirmesi (özet, atıf, konular)
Özet
OpenAlex · İngilizce
We introduce a highly heterogeneous and programmable compute-in-memory (CIM) accelerator architecture for deep neural network (DNN) inference. This architecture combines spatially distributed CIM memory array “tiles” for weight-stationary, energy-efficient multiply–accumulate (MAC) operations, together with heterogeneous special-function compute cores for auxiliary digital computation. Massively parallel vectors of neuron activation data are exchanged over short distances using a dense and efficient circuit-switched 2-D mesh, offering full end-to-end support for a wide range of DNN workloads, including CNNs, long-short-term-memory (LSTM), and transformers. We discuss the design of the “analog fabric”—the 2-D grid of tiles and compute cores interconnected by the 2-D mesh—and address the efficiency in both mapping of DNNs onto the hardware and in pipelining of various DNN workloads across a range of batch sizes. We show, for the first time, system-level assessments using projected component parameters for a realistic “analog AI” system, based on dense crossbar arrays of low-power nonvolatile analog memory elements, while incorporating a single common analog fabric design that can scale to large networks by introducing data transport between multiple analog AI chips. Our performance estimates for several networks, including large LSTM and bidirectional encoder representations from transformers (BERT), show highly competitive throughput while offering$40\times $–$140\times $higher energy efficiency than NVIDIA A100—thus illustrating the strong promise of analog AI and the proposed architecture for DNN inference applications.
Konular
Atıflar
OpenAlex cited_by_count. WoS veya Scopus atıf sayısı değildir; o kaynaklar için ayrı kolon yoktur.
74 atıf
OpenAlex cited_by_count (önbellek / veritabanı)
Yerel katalogda bu makaleye atıf yapan 17 yayın (OpenAlex referans eşleşmesi; tam dünya listesi değildir).
- Recent Advances and Future Prospects for Memristive Materials, Devices, and Systems 2023
- Efficient scaling of large language models with mixture of experts and 3D analog in-memory computing 2025
- Heterogeneous Embedded Neural Processing Units Utilizing PCM-Based Analog In-Memory Computing 2024
- AnalogNAS: A Neural Network Design Framework for Accurate Inference with Analog In-Memory Computing 2023
- CiMBA: Accelerating Genome Sequencing Through On-Device Basecalling via Compute-in-Memory 2025
- Design of Analog-AI Hardware Accelerators for Transformer-based Language Models (Invited) 2023
- Impact of Phase-Change Memory Drift on Energy Efficiency and Accuracy of Analog Compute-in-Memory Deep Learning Inference (Invited) 2023
- Deep learning software stacks for analogue in-memory computing-based accelerators 2025
- A Precision-Optimized Fixed-Point Near-Memory Digital Processing Unit for Analog In-Memory Computing 2024
- Analog AI Accelerators for Transformer-based Language Models: Hardware, Workload, and Power Performance 2025