Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Oktay Günlük

Generalized Linear Rule Models

Jun 05, 2019

Dennis Wei, Sanjeeb Dash, Tian Gao, Oktay Günlük

Figure 1 for Generalized Linear Rule Models

Figure 2 for Generalized Linear Rule Models

Figure 3 for Generalized Linear Rule Models

Figure 4 for Generalized Linear Rule Models

Abstract:This paper considers generalized linear models using rule-based features, also referred to as rule ensembles, for regression and probabilistic classification. Rules facilitate model interpretation while also capturing nonlinear dependences and interactions. Our problem formulation accordingly trades off rule set complexity and prediction accuracy. Column generation is used to optimize over an exponentially large space of rules without pre-generating a large subset of candidates or greedily boosting rules one by one. The column generation subproblem is solved using either integer programming or a heuristic optimizing the same objective. In experiments involving logistic and linear regression, the proposed methods obtain better accuracy-complexity trade-offs than existing rule ensemble algorithms. At one end of the trade-off, the methods are competitive with less interpretable benchmark models.

* Published in the Proceedings of the 36th International Conference on Machine Learning (ICML), PMLR 97:6687-6696, 2019. 17 pages, 7 figures

Via

Access Paper or Ask Questions

Boolean Decision Rules via Column Generation

May 24, 2018

Sanjeeb Dash, Oktay Günlük, Dennis Wei

Figure 1 for Boolean Decision Rules via Column Generation

Figure 2 for Boolean Decision Rules via Column Generation

Figure 3 for Boolean Decision Rules via Column Generation

Figure 4 for Boolean Decision Rules via Column Generation

Abstract:This paper considers the learning of Boolean rules in either disjunctive normal form (DNF, OR-of-ANDs, equivalent to decision rule sets) or conjunctive normal form (CNF, AND-of-ORs) as an interpretable model for classification. An integer program is formulated to optimally trade classification accuracy for rule simplicity. Column generation (CG) is used to efficiently search over an exponential number of candidate clauses (conjunctions or disjunctions) without the need for heuristic rule mining. This approach also bounds the gap between the selected rule set and the best possible rule set on the training data. To handle large datasets, we propose an approximate CG algorithm using randomization. Compared to three recently proposed alternatives, the CG algorithm dominates the accuracy-simplicity trade-off in 7 out of 15 datasets. When maximized for accuracy, CG is competitive with rule learners designed for this purpose, sometimes finding significantly simpler solutions that are no less accurate.

Via

Access Paper or Ask Questions