Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:TResNet: High Performance GPU-Dedicated Architecture

Mar 30, 2020

Tal Ridnik, Hussam Lawen, Asaf Noy, Itamar Friedman

Figure 1 for TResNet: High Performance GPU-Dedicated Architecture

Figure 2 for TResNet: High Performance GPU-Dedicated Architecture

Figure 3 for TResNet: High Performance GPU-Dedicated Architecture

Figure 4 for TResNet: High Performance GPU-Dedicated Architecture

Share this with someone who'll enjoy it:

Abstract:Many deep learning models, developed in recent years, reach higher ImageNet accuracy than ResNet50, with fewer or comparable FLOPS count. While FLOPs are often seen as a proxy for network efficiency, when measuring actual GPU training and inference throughput, vanilla ResNet50 is usually significantly faster than its recent competitors, offering better throughput-accuracy trade-off. In this work, we introduce a series of architecture modifications that aim to boost neural networks' accuracy, while retaining their GPU training and inference efficiency. We first demonstrate and discuss the bottlenecks induced by FLOPs-optimizations. We then suggest alternative designs that better utilize GPU structure and assets. Finally, we introduce a new family of GPU-dedicated models, called TResNet, which achieve better accuracy and efficiency than previous ConvNets. Using a TResNet model, with similar GPU throughput to ResNet50, we reach 80.7% top-1 accuracy on ImageNet. Our TResNet models also transfer well and achieve state-of-the-art accuracy on competitive datasets such as Stanford cars (96.0%), CIFAR-10 (99.0%), CIFAR-100 (91.5%) and Oxford-Flowers (99.1%). Implementation is available at: https://github.com/mrT23/TResNet

* 9 pages, 5 figures

View paper on

Share this with someone who'll enjoy it:

Title:TResNet: High Performance GPU-Dedicated Architecture

Paper and Code