Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Neural Tangent Kernel Eigenvalues Accurately Predict Generalization

Oct 13, 2021

James B. Simon, Madeline Dickens, Michael R. DeWeese

Figure 1 for Neural Tangent Kernel Eigenvalues Accurately Predict Generalization

Figure 2 for Neural Tangent Kernel Eigenvalues Accurately Predict Generalization

Figure 3 for Neural Tangent Kernel Eigenvalues Accurately Predict Generalization

Figure 4 for Neural Tangent Kernel Eigenvalues Accurately Predict Generalization

Share this with someone who'll enjoy it:

Abstract:Finding a quantitative theory of neural network generalization has long been a central goal of deep learning research. We extend recent results to demonstrate that, by examining the eigensystem of a neural network's "neural tangent kernel", one can predict its generalization performance when learning arbitrary functions. Our theory accurately predicts not only test mean-squared-error but all first- and second-order statistics of the network's learned function. Furthermore, using a measure quantifying the "learnability" of a given target function, we prove a new "no-free-lunch" theorem characterizing a fundamental tradeoff in the inductive bias of wide neural networks: improving a network's generalization for a given target function must worsen its generalization for orthogonal functions. We further demonstrate the utility of our theory by analytically predicting two surprising phenomena - worse-than-chance generalization on hard-to-learn functions and nonmonotonic error curves in the small data regime - which we subsequently observe in experiments. Though our theory is derived for infinite-width architectures, we find it agrees with networks as narrow as width 20, suggesting it is predictive of generalization in practical neural networks. Code replicating our results is available at https://github.com/james-simon/eigenlearning.

* 10 pages (main text), 24 pages (total), 10 figures

View paper on

OpenReview

Share this with someone who'll enjoy it:

Title:Neural Tangent Kernel Eigenvalues Accurately Predict Generalization

Paper and Code