Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?

Aug 21, 2024

Francesco Innocenti, El Mehdi Achour, Ryan Singh, Christopher L. Buckley

Figure 1 for Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?

Figure 2 for Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?

Figure 3 for Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?

Figure 4 for Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?

Share this with someone who'll enjoy it:

Abstract:Predictive coding (PC) is an energy-based learning algorithm that performs iterative inference over network activities before weight updates. Recent work suggests that PC can converge in fewer learning steps than backpropagation thanks to its inference procedure. However, these advantages are not always observed, and the impact of PC inference on learning is theoretically not well understood. Here, we study the geometry of the PC energy landscape at the (inference) equilibrium of the network activities. For deep linear networks, we first show that the equilibrated energy is simply a rescaled mean squared error loss with a weight-dependent rescaling. We then prove that many highly degenerate (non-strict) saddles of the loss including the origin become much easier to escape (strict) in the equilibrated energy. Our theory is validated by experiments on both linear and non-linear networks. Based on these results, we conjecture that all the saddles of the equilibrated energy are strict. Overall, this work suggests that PC inference makes the loss landscape more benign and robust to vanishing gradients, while also highlighting the challenge of speeding up PC inference on large-scale models.

* 26 pages, 12 figures

View paper on

Share this with someone who'll enjoy it:

Title:Only Strict Saddles in the Energy Landscape of Predictive Coding Networks?

Paper and Code