A fine-grained mixed precision DNN accelerator using a two-stage big–little core RISC-V MCU

2Citations
Citations of this article
9Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Deep neural networks (DNNs) are widely used in modern AI systems, and their dedicated accelerators have become a promising option for edge scenarios due to the energy efficiency and high performance. Since the DNN model requires significant storage and computation resources, various energy efficient DNN accelerators or algorithms have been proposed for edge devices. Many quantization algorithms for efficient DNN training have been proposed, where the weights in the DNN layers are quantified to small/zero values, therefore requiring much fewer effective bits, i.e., low-precision bits in both arithmetic and storage units. Such sparsity can be leveraged to remove non-effective bits and reduce the design cost, however at the cost of accuracy degradation, as some key operations still demand a higher precision. Therefore, in this paper, we propose a universal mixed-precision DNN accelerator architecture that can simultaneously support mixed-precision DNN arithmetic operations. A big–little core controller based on RISC-V is implemented to effectively control the datapath, and assign the arithmetic operations to full precision and low precision process units, respectively. Experimental results show that, with the proposed designs, we can save 16% chip area and 45.2% DRAM access compared with the state-of-the-art design.

Cite

CITATION STYLE

APA

Zhang, L., Lv, Q., Gao, D., Zhou, X., Meng, W., Yang, Q., & Zhuo, C. (2023). A fine-grained mixed precision DNN accelerator using a two-stage big–little core RISC-V MCU. Integration, 88, 241–248. https://doi.org/10.1016/j.vlsi.2022.10.006

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free