Loss-optimal classification trees: a generalized framework and the logistic case

被引:0
|
作者
Aldinucci, Tommaso [1 ]
Lapucci, Matteo [1 ]
机构
[1] Univ Florence, Dipartimento Ingn Informaz, Via Santa Marta 3, I-50139 Florence, Italy
关键词
Optimal classification trees; Logistic regression; Interpretability; Mixed-integer programming; DECISION TREES; REGRESSION; SELECTION; OPTIMIZATION; INDUCTION; MODELS;
D O I
10.1007/s11750-024-00674-y
中图分类号
C93 [管理学]; O22 [运筹学];
学科分类号
070105 ; 12 ; 1201 ; 1202 ; 120202 ;
摘要
Classification trees are one of the most common models in interpretable machine learning. Although such models are usually built with greedy strategies, in recent years, thanks to remarkable advances in mixed-integer programming (MIP) solvers, several exact formulations of the learning problem have been developed. In this paper, we argue that some of the most relevant ones among these training models can be encapsulated within a general framework, whose instances are shaped by the specification of loss functions and regularizers. Next, we introduce a novel realization of this framework: specifically, we consider the logistic loss, handled in the MIP setting by a piece-wise linear approximation, and couple it with & ell; 1 \documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\ell _1$$\end{document} -regularization terms. The resulting optimal logistic classification tree model numerically proves to be able to induce trees with enhanced interpretability properties and competitive generalization capabilities, compared to the state-of-the-art MIP-based approaches.
引用
收藏
页码:323 / 350
页数:28
相关论文
共 50 条
  • [1] Optimal generalized logistic estimator
    Varathan, Nagarajah
    Wijekoon, Pushpakanthie
    COMMUNICATIONS IN STATISTICS-THEORY AND METHODS, 2018, 47 (02) : 463 - 474
  • [2] Optimal classification trees
    Dimitris Bertsimas
    Jack Dunn
    Machine Learning, 2017, 106 : 1039 - 1082
  • [3] Optimal classification trees
    Bertsimas, Dimitris
    Dunn, Jack
    MACHINE LEARNING, 2017, 106 (07) : 1039 - 1082
  • [4] LOSS-OPTIMAL PWM WAVEFORMS FOR VARIABLE-SPEED INDUCTION-MOTOR DRIVES
    DEBUCK, F
    GISTELINCK, P
    DEBACKER, D
    IEE PROCEEDINGS-B ELECTRIC POWER APPLICATIONS, 1983, 130 (05): : 310 - 320
  • [5] Optimal randomized classification trees
    Blanquero, Rafael
    Carrizosa, Emilio
    Molero-Rio, Cristina
    Morales, Dolores Romero
    COMPUTERS & OPERATIONS RESEARCH, 2021, 132
  • [6] Strong Optimal Classification Trees
    Aghaei, Sina
    Gomez, Andres
    Vayanos, Phebe
    OPERATIONS RESEARCH, 2024,
  • [7] Margin optimal classification trees
    D'Onofrio, Federico
    Grani, Giorgio
    Monaci, Marta
    Palagi, Laura
    COMPUTERS & OPERATIONS RESEARCH, 2024, 161
  • [8] OPTIMAL MULTIWAY GENERALIZED SPLIT TREES
    CHEN, GH
    LIU, LT
    INTERNATIONAL JOURNAL OF COMPUTER MATHEMATICS, 1991, 41 (1-2) : 39 - 47
  • [9] Generalized neural trees for pattern classification
    Foresti, GL
    Micheloni, C
    IEEE TRANSACTIONS ON NEURAL NETWORKS, 2002, 13 (06): : 1540 - 1547
  • [10] The existence of optimal parameters of the generalized logistic function
    Jukic, D
    Scitovski, R
    APPLIED MATHEMATICS AND COMPUTATION, 1996, 77 (2-3) : 281 - 294