Abstract
No framework exists that can explain and predict the generalisation ability of deep neural networks in general circumstances. In fact, this question has not been answered for some of the least complicated of neural network architectures: fully-connected feedforward networks with rectified linear activations and a limited number of hidden layers. For such an architecture, we show how adding a summary layer to the network makes it more amenable to analysis, and allows us to define the conditions that are required to guarantee that a set of samples will all be classified correctly. This process does not describe the generalisation behaviour of these networks, but produces a number of metrics that are useful for probing their learning and generalisation behaviour. We support the analytical conclusions with empirical results, both to confirm that the mathematical guarantees hold in practice, and to demonstrate the use of the analysis process.
Author supplied keywords
Cite
CITATION STYLE
Davel, M. H. (2020). Using Summary Layers to Probe Neural Network Behaviour. South African Computer Journal, 32(2), 102–123. https://doi.org/10.18489/SACJ.V32I2.861
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.