Item request has been placed! ×
Item request cannot be made. ×
loading  Processing Request

GPGPU Accelerated Deep Object Classification on a Heterogeneous Mobile Platform

Item request has been placed! ×
Item request cannot be made. ×
loading   Processing Request
  • معلومة اضافية
    • Contributors:
      Rizvi, SYED TAHIR HUSSAIN; Cabodi, Gianpiero; Patti, Deni; Francini, Gianluca
    • بيانات النشر:
      MDPI
    • الموضوع:
      2016
    • Collection:
      PORTO@iris (Publications Open Repository TOrino - Politecnico di Torino)
    • نبذة مختصرة :
      Deep convolutional neural networks achieve state-of-the-art performance in image classification. The computational and memory requirements of such networks are however huge, and that is an issue on embedded devices due to their constraints. Most of this complexity derives from the convolutional layers and in particular from the matrix multiplications they entail. This paper proposes a complete approach to image classification providing common layers used in neural networks. Namely, the proposed approach relies on a heterogeneous CPU-GPU scheme for performing convolutions in the transform domain. The Compute Unified Device Architecture(CUDA)-based implementation of the proposed approach is evaluated over three different image classification networks on a Tegra K1 CPU-GPU mobile processor. Experiments show that the presented heterogeneous scheme boasts a 50 speedup over the CPU-only reference and outperforms a GPU-based reference by 2, while slashing the power consumption by nearly 30%.
    • Relation:
      info:eu-repo/semantics/altIdentifier/wos/WOS:000392387600010; volume:5; issue:4; numberofpages:17; journal:ELECTRONICS; http://hdl.handle.net/11583/2659082; info:eu-repo/semantics/altIdentifier/scopus/2-s2.0-85027446278
    • الرقم المعرف:
      10.3390/electronics5040088
    • الدخول الالكتروني :
      http://hdl.handle.net/11583/2659082
      https://doi.org/10.3390/electronics5040088
    • Rights:
      info:eu-repo/semantics/openAccess
    • الرقم المعرف:
      edsbas.6ED6438D