Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:PaloBoost: An Overfitting-robust TreeBoost with Out-of-Bag Sample Regularization Techniques

Jul 22, 2018

Yubin Park, Joyce C. Ho

Figure 1 for PaloBoost: An Overfitting-robust TreeBoost with Out-of-Bag Sample Regularization Techniques

Figure 2 for PaloBoost: An Overfitting-robust TreeBoost with Out-of-Bag Sample Regularization Techniques

Figure 3 for PaloBoost: An Overfitting-robust TreeBoost with Out-of-Bag Sample Regularization Techniques

Figure 4 for PaloBoost: An Overfitting-robust TreeBoost with Out-of-Bag Sample Regularization Techniques

Share this with someone who'll enjoy it:

Abstract:Stochastic Gradient TreeBoost is often found in many winning solutions in public data science challenges. Unfortunately, the best performance requires extensive parameter tuning and can be prone to overfitting. We propose PaloBoost, a Stochastic Gradient TreeBoost model that uses novel regularization techniques to guard against overfitting and is robust to parameter settings. PaloBoost uses the under-utilized out-of-bag samples to perform gradient-aware pruning and estimate adaptive learning rates. Unlike other Stochastic Gradient TreeBoost models that use the out-of-bag samples to estimate test errors, PaloBoost treats the samples as a second batch of training samples to prune the trees and adjust the learning rates. As a result, PaloBoost can dynamically adjust tree depths and learning rates to achieve faster learning at the start and slower learning as the algorithm converges. We illustrate how these regularization techniques can be efficiently implemented and propose a new formula for calculating feature importance to reflect the node coverages and learning rates. Extensive experimental results on seven datasets demonstrate that PaloBoost is robust to overfitting, is less sensitivity to the parameters, and can also effectively identify meaningful features.

View paper on

Share this with someone who'll enjoy it:

Title:PaloBoost: An Overfitting-robust TreeBoost with Out-of-Bag Sample Regularization Techniques

Paper and Code