Performance assessment of deep learning frameworks through metrics of CPU hardware exploitation on an embedded platform

2Citations
Citations of this article
8Readers
Mendeley users who have this article in their library.

Abstract

In this paper, we analyze heterogeneous performance exhibited by some popular deep learning software frameworks for visual inference on a resource-constrained hardware platform. Benchmarking of Caffe, OpenCV, TensorFlow, and Caffe2 is performed on the same set of convolutional neural networks in terms of instantaneous throughput, power consumption, memory footprint, and CPU utilization. To understand the resulting dissimilar behavior, we thoroughly examine how the resources in the processor are differently exploited by these frameworks. We demonstrate that a strong correlation exists between hardware events occurring in the processor and inference performance. The proposedhardware-aware analysis aims to findlimitations andbottlenecks emerging from the jointinteraction offrameworks andnetworks on a particular CPU-based platform. This provides insight into introducing suitable modifications in bothtypes of components to enhance their global performance. It also facilitates the selection of frameworks and networks among a large diversity of these components available these days for visual understanding.

Cite

CITATION STYLE

APA

Velasco-Montero, D., Fernández-Berni, J., Carmona-Galán, R., & Rodríguez-Vázquez, Á. (2020). Performance assessment of deep learning frameworks through metrics of CPU hardware exploitation on an embedded platform. International Journal of Electrical and Computer Engineering Systems, 11(1), 1–11. https://doi.org/10.32985/ijeces.11.1.1

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free