Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:An Image Patch is a Wave: Phase-Aware Vision MLP

Nov 25, 2021

Yehui Tang, Kai Han, Jianyuan Guo, Chang Xu, Yanxi Li, Chao Xu, Yunhe Wang

Figure 1 for An Image Patch is a Wave: Phase-Aware Vision MLP

Figure 2 for An Image Patch is a Wave: Phase-Aware Vision MLP

Figure 3 for An Image Patch is a Wave: Phase-Aware Vision MLP

Figure 4 for An Image Patch is a Wave: Phase-Aware Vision MLP

Share this with someone who'll enjoy it:

Abstract:Different from traditional convolutional neural network (CNN) and vision transformer, the multilayer perceptron (MLP) is a new kind of vision model with extremely simple architecture that only stacked by fully-connected layers. An input image of vision MLP is usually split into multiple tokens (patches), while the existing MLP models directly aggregate them with fixed weights, neglecting the varying semantic information of tokens from different images. To dynamically aggregate tokens, we propose to represent each token as a wave function with two parts, amplitude and phase. Amplitude is the original feature and the phase term is a complex value changing according to the semantic contents of input images. Introducing the phase term can dynamically modulate the relationship between tokens and fixed weights in MLP. Based on the wave-like token representation, we establish a novel Wave-MLP architecture for vision tasks. Extensive experiments demonstrate that the proposed Wave-MLP is superior to the state-of-the-art MLP architectures on various vision tasks such as image classification, object detection and semantic segmentation.

View paper on

Share this with someone who'll enjoy it:

Title:An Image Patch is a Wave: Phase-Aware Vision MLP

Paper and Code