Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Measuring the Effects of Data Parallelism on Neural Network Training

Nov 21, 2018

Christopher J. Shallue, Jaehoon Lee, Joseph Antognini, Jascha Sohl-Dickstein, Roy Frostig, George E. Dahl

Figure 1 for Measuring the Effects of Data Parallelism on Neural Network Training

Figure 2 for Measuring the Effects of Data Parallelism on Neural Network Training

Figure 3 for Measuring the Effects of Data Parallelism on Neural Network Training

Figure 4 for Measuring the Effects of Data Parallelism on Neural Network Training

Share this with someone who'll enjoy it:

Abstract:Recent hardware developments have made unprecedented amounts of data parallelism available for accelerating neural network training. Among the simplest ways to harness next-generation accelerators is to increase the batch size in standard mini-batch neural network training algorithms. In this work, we aim to experimentally characterize the effects of increasing the batch size on training time, as measured in the number of steps necessary to reach a goal out-of-sample error. Eventually, increasing the batch size will no longer reduce the number of training steps required, but the exact relationship between the batch size and how many training steps are necessary is of critical importance to practitioners, researchers, and hardware designers alike. We study how this relationship varies with the training algorithm, model, and data set and find extremely large variation between workloads. Along the way, we reconcile disagreements in the literature on whether batch size affects model quality. Finally, we discuss the implications of our results for efforts to train neural networks much faster in the future.

* Submitted to JMLR

View paper on

Share this with someone who'll enjoy it:

Title:Measuring the Effects of Data Parallelism on Neural Network Training

Paper and Code