Dynamic Quantization Range Control for Analog-in-Memory Neural Networks Acceleration

Published: 06 June 2022 Publication History


Analog in Memory Computing (AiMC) based neural network acceleration is a promising solution to increase the energy efficiency of deep neural networks deployment. However, the quantization requirements of these analog systems are not compatible with state-of-the-art neural network quantization techniques. Indeed, while the quantization of the weights and activations is considered by modern deep neural network quantization techniques, AiMC accelerators also impose the quantization of each Matrix Vector Multiplication (MVM) result. In most demonstrated AiMC implementations, the quantization range of MVM results is considered a fixed parameter of the accelerator. This work demonstrates that dynamic control over this quantization range is possible but also desirable for analog neural networks acceleration. An AiMC compatible quantization flow coupled with a hardware aware quantization range driving technique is introduced to fully exploit these dynamic ranges. Using CIFAR-10 and ImageNet as benchmarks, the proposed solution results in networks that are both more accurate and more robust to the inherent vulnerability of analog circuits than fixed quantization range based approaches.
  Input-Conditioned Quantisation for ENOB Improvement in CIM ADC Columns Targeting Large-Length Partial Sums, IEEE Transactions on Circuits and Systems II: Express Briefs, 2024
  Measurement-driven neural-network training for integrated magnetic tunnel junction arrays, Physical Review Applied, 2024
  Modelling and Optimization of a Mixed-Signal Accelerator for Deep Neural Networks, 2023 19th International Conference on Synthesis, Modeling, Analysis and Simulation Methods and Applications to Circuit Design (SMACD), 2023
ACM Transactions on Design Automation of Electronic Systems  Volume 27, Issue 5
September 2022
Publication History

Published: 06 June 2022
Online AM: 24 February 2022
Accepted: 01 November 2021
Revised: 01 October 2021
Received: 01 July 2021
Published in TODAES Volume 27, Issue 5


Author Tags

  1. Neural networks
  2. quantization
  3. in-memory-computing


Funding Sources

  • Flemish Government (AI Research Program) and KU Leuven


