Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Learning Energy Networks with Generalized Fenchel-Young Losses

May 19, 2022

Mathieu Blondel, Felipe Llinares-López, Robert Dadashi, Léonard Hussenot, Matthieu Geist

Figure 1 for Learning Energy Networks with Generalized Fenchel-Young Losses

Figure 2 for Learning Energy Networks with Generalized Fenchel-Young Losses

Figure 3 for Learning Energy Networks with Generalized Fenchel-Young Losses

Figure 4 for Learning Energy Networks with Generalized Fenchel-Young Losses

Share this with someone who'll enjoy it:

Abstract:Energy-based models, a.k.a. energy networks, perform inference by optimizing an energy function, typically parametrized by a neural network. This allows one to capture potentially complex relationships between inputs and outputs. To learn the parameters of the energy function, the solution to that optimization problem is typically fed into a loss function. The key challenge for training energy networks lies in computing loss gradients, as this typically requires argmin/argmax differentiation. In this paper, building upon a generalized notion of conjugate function, which replaces the usual bilinear pairing with a general energy function, we propose generalized Fenchel-Young losses, a natural loss construction for learning energy networks. Our losses enjoy many desirable properties and their gradients can be computed efficiently without argmin/argmax differentiation. We also prove the calibration of their excess risk in the case of linear-concave energies. We demonstrate our losses on multilabel classification and imitation learning tasks.

View paper on

OpenReview

Share this with someone who'll enjoy it:

Title:Learning Energy Networks with Generalized Fenchel-Young Losses

Paper and Code