Update Disturbance-Resilient Analog ReRAM Crossbar Arrays for In-Memory Deep Learning Accelerators
Wooseok Choi, Tommaso Stecconi, Donato Francesco Falcone, Matteo Galetta, Victoria Clerico, Elisa Zaccaria, Mamidala Saketh Ram, Antonio La Porta, Folkert Horst, Daniel Jubin, Matias Senger, Marilyne Sousa, Steffen Reidt, Ralph Heller, Bernabe Linares-Barranco, Valeria Bragaglia, Bert Jan Offrein
Abstract
Resistive memory (ReRAM) technologies with crossbar array architectures hold significant potential for analog AI accelerator hardware, enabling both in-memory inference and training. Recent developments have successfully demonstrated inference acceleration by offloading compute-heavy training workloads to off-chip digital processors. However, in-memory acceleration of training algorithms is crucial for more sustainable and power-efficient AI, but still in an early stage of research. This study addresses in-memory training acceleration using analog ReRAM arrays, focusing on a key challenge during fully parallel weight updates: disturbances of the weight values in cross-point devices. A ReRAM device solution is presented on 350 nm silicon technology, utilizing a resistive switching conductive metal oxide (CMO) formed on a nanoscale conductive filament within a HfOx layer. The devices not only exhibit 60 ns fast, non-volatile analog switching, but also demonstrates outstanding resilience to update disturbances, enduring over 100k pulses. The disturbance tolerance of the ReRAM is analyzed using COMSOL Multiphysics simulations, modeling the filament-induced thermoelectric energy concentration that results in a highly nonlinear device responses to input voltage amplitudes. Disturbance-free parallel weight mapping is also demonstrated on the back-end-of-line integrated ReRAM array chip. Finally, comprehensive hardware-aware neural network simulations validate the potential of our ReRAM for in-memory deep learning accelerators capable of fully parallel weight updates.
Create a lesson
Related papers
Co-occurrence-Aware Quadratic Assignment for Local Feature Matching in Simultaneous Localization and Mapping
Yutaka Yamada, Yohei Hamakawa, Yutaro Ishigaki et al.
Composability rather than computation sets the cost of an analog EML hardware fabric
C. Teuscher
A Hybrid Quantum-Classical Coordination Architecture for Portfolio Optimization via Global Context Injection
Xiaoguang Yang, Menghan Dou, Guoping Guo
A Game-Theoretic Framework for Incentive-Compatible AI training Under Renewable-Energy Constraints
Konstantinos Varsos, Ramin Khalili, Adamantia Stamou et al.
System-Technology Co-Evaluation of A7 CFET and A10 NSFET Technologies from Cell Parasitics to Chip Reliability
Mahdi Benkhelifa, Leon Mayr, Hadi Nour Eddine et al.
Feasibility and Memory Mechanisms of Chern-Simons Context Reservoir Computation
Jyotiranjan Beuria, Venkatesh H. Chembrolu