Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:ActiveMLP: An MLP-like Architecture with Active Token Mixer

Mar 11, 2022

Guoqiang Wei, Zhizheng Zhang, Cuiling Lan, Yan Lu, Zhibo Chen

Figure 1 for ActiveMLP: An MLP-like Architecture with Active Token Mixer

Figure 2 for ActiveMLP: An MLP-like Architecture with Active Token Mixer

Figure 3 for ActiveMLP: An MLP-like Architecture with Active Token Mixer

Figure 4 for ActiveMLP: An MLP-like Architecture with Active Token Mixer

Share this with someone who'll enjoy it:

Abstract:This paper presents ActiveMLP, a general MLP-like backbone for computer vision. The three existing dominant network families, i.e., CNNs, Transformers and MLPs, differ from each other mainly in the ways to fuse contextual information into a given token, leaving the design of more effective token-mixing mechanisms at the core of backbone architecture development. In ActiveMLP, we propose an innovative token-mixer, dubbed Active Token Mixer (ATM), to actively incorporate contextual information from other tokens in the global scope into the given one. This fundamental operator actively predicts where to capture useful contexts and learns how to fuse the captured contexts with the original information of the given token at channel levels. In this way, the spatial range of token-mixing is expanded and the way of token-mixing is reformed. With this design, ActiveMLP is endowed with the merits of global receptive fields and more flexible content-adaptive information fusion. Extensive experiments demonstrate that ActiveMLP is generally applicable and comprehensively surpasses different families of SOTA vision backbones by a clear margin on a broad range of vision tasks, including visual recognition and dense prediction tasks. The code and models will be available at https://github.com/microsoft/ActiveMLP.

View paper on

Share this with someone who'll enjoy it:

Title:ActiveMLP: An MLP-like Architecture with Active Token Mixer

Paper and Code