DNNZip: Selective Layers Compression Technique in Deep Neural Network Accelerators - ETIS, équipe ASTRE
Communication Dans Un Congrès Année : 2020

DNNZip: Selective Layers Compression Technique in Deep Neural Network Accelerators

Habiba Lahdhiri
  • Fonction : Auteur
  • PersonId : 1062534
Davide Patti
  • Fonction : Auteur
  • PersonId : 1075156
Giuseppe Ascia
  • Fonction : Auteur
  • PersonId : 1075157
Emmanuelle Bourdel
  • Fonction : Auteur
  • PersonId : 884035
Vincenzo Catania
  • Fonction : Auteur
  • PersonId : 1075158

Résumé

In Deep Neural Network (DNN) accelerators, the on-chip traffic and memory traffic accounts for a relevant fraction of the inference latency and energy consumption. A major component of such traffic is due to the moving of the DNN model parameters from the main memory to the memory interface and from the latter to the processing elements (PEs) of the accelerator. In this paper, we present DNNZip, a technique aimed at compressing the model parameters of a DNN, thus resulting in significant energy and performance improvement. DNNZip implements a lossy compression whose compression ratio is tuned based on the maximum tolerated error on the model parameters provided by the user. DNNZip is assessed on several convolutional NNs and the trade-off inference energy saving vs. inference latency reduction vs. network accuracy degradation is discussed. We found that up to 64% energy saving, and up to 67% latency reduction can be obtained with a limited impact on the accuracy of the network.
Fichier principal
Vignette du fichier
DNNZip_DSD20.pdf (540.72 Ko) Télécharger le fichier
Origine Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-02906973 , version 1 (26-07-2020)

Identifiants

  • HAL Id : hal-02906973 , version 1

Citer

Habiba Lahdhiri, Maurizio Palesi, Salvatore Monteleone, Davide Patti, Giuseppe Ascia, et al.. DNNZip: Selective Layers Compression Technique in Deep Neural Network Accelerators. Euromicro Conference on Digital System Design DSD, Aug 2020, Portorož, Slovenia. ⟨hal-02906973⟩
211 Consultations
312 Téléchargements

Partager

More