A law of data separation in deep learning

N/ACitations
Citations of this article
35Readers
Mendeley users who have this article in their library.
Get full text

Abstract

While deep learning has enabled significant advances in many areas of science, its blackbox nature hinders architecture design for future artificial intelligence applications and interpretation for high-stakes decision-makings. We addressed this issue by studying the fundamental question of how deep neural networks process data in the intermediate layers. Our finding is a simple and quantitative law that governs how deep neural networks separate data according to class membership throughout all layers for classification. This law shows that each layer improves data separation at a constant geometric rate, and its emergence is observed in a collection of network architectures and datasets during training. This law offers practical guidelines for designing architectures, improving model robustness and out-of-sample performance, as well as interpreting the predictions.

Cite

CITATION STYLE

APA

He, H., & Su, W. J. (2023). A law of data separation in deep learning. Proceedings of the National Academy of Sciences of the United States of America, 120(36). https://doi.org/10.1073/pnas.2221704120

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free