Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Conditional computation in neural networks: principles and research trends

Mar 12, 2024

Simone Scardapane, Alessandro Baiocchi, Alessio Devoto, Valerio Marsocci, Pasquale Minervini, Jary Pomponi

Figure 1 for Conditional computation in neural networks: principles and research trends

Figure 2 for Conditional computation in neural networks: principles and research trends

Figure 3 for Conditional computation in neural networks: principles and research trends

Figure 4 for Conditional computation in neural networks: principles and research trends

Share this with someone who'll enjoy it:

Abstract:This article summarizes principles and ideas from the emerging area of applying \textit{conditional computation} methods to the design of neural networks. In particular, we focus on neural networks that can dynamically activate or de-activate parts of their computational graph conditionally on their input. Examples include the dynamic selection of, e.g., input tokens, layers (or sets of layers), and sub-modules inside each layer (e.g., channels in a convolutional filter). We first provide a general formalism to describe these techniques in an uniform way. Then, we introduce three notable implementations of these principles: mixture-of-experts (MoEs) networks, token selection mechanisms, and early-exit neural networks. The paper aims to provide a tutorial-like introduction to this growing field. To this end, we analyze the benefits of these modular designs in terms of efficiency, explainability, and transfer learning, with a focus on emerging applicative areas ranging from automated scientific discovery to semantic communication.

* Under review at Intelligenza Artificiale (IOS Press)

View paper on

Share this with someone who'll enjoy it:

Title:Conditional computation in neural networks: principles and research trends

Paper and Code