A novel estimator for the two-way partial AUC

被引:1
|
作者
Neto, Elias Chaibub [1 ]
Yadav, Vijay [1 ]
Sieberts, Solveig K. [1 ]
Omberg, Larsson [1 ]
机构
[1] Sage Bionetworks, 2901 Third Ave, Seattle, WA 98121 USA
关键词
ROC curve; AUC; Partial AUC; Diagnostic testing; Machine learning performance metric; COMPUTER-AIDED DIAGNOSIS; PARTIAL AREA; ROC; METHODOLOGY;
D O I
10.1186/s12911-023-02382-2
中图分类号
R-058 [];
学科分类号
摘要
Background The two-way partial AUC has been recently proposed as a way to directly quantify partial area under the ROC curve with simultaneous restrictions on the sensitivity and specificity ranges of diagnostic tests or classifiers. The metric, as originally implemented in the tpAUC R package, is estimated using a nonparametric estimator based on a trimmed Mann-Whitney U-statistic, which becomes computationally expensive in large sample sizes. (Its computational complexity is of order O(n(x)n(y)), where n(x) and n(y) represent the number of positive and negative cases, respectively). This is problematic since the statistical methodology for comparing estimates generated from alternative diagnostic tests/classifiers relies on bootstrapping resampling and requires repeated computations of the estimator on a large number of bootstrap samples. Methods By leveraging the graphical and probabilistic representations of the AUC, partial AUCs, and two-way partial AUC, we derive a novel estimator for the two-way partial AUC, which can be directly computed from the output of any software able to compute AUC and partial AUCs. We implemented our estimator using the computationally efficient pROC R package, which leverages a nonparametric approach using the trapezoidal rule for the computation of AUC and partial AUC scores. (Its computational complexity is of order O(n log n), where n = n(x) + n(y).). We compare the empirical bias and computation time of the proposed estimator against the original estimator provided in the tpAUC package in a series of simulation studies and on two real datasets. Results Our estimator tended to be less biased than the original estimator based on the trimmed Mann-Whitney U-statistic across all experiments (and showed considerably less bias in the experiments based on small sample sizes). But, most importantly, because the computational complexity of the proposed estimator is of order O(n log n), rather than O(nxny), it is much faster to compute when sample sizes are large. Conclusions The proposed estimator provides an improvement for the computation of two-way partial AUC, and allows the comparison of diagnostic tests/machine learning classifiers in large datasets where repeated computations of the original estimator on bootstrap samples become too expensive to compute.
引用
收藏
页数:17
相关论文
共 50 条
  • [1] A novel estimator for the two-way partial AUC
    Elias Chaibub Neto
    Vijay Yadav
    Solveig K. Sieberts
    Larsson Omberg
    BMC Medical Informatics and Decision Making, 24
  • [2] Two-way partial AUC and its properties
    Yang, Hanfang
    Lu, Kun
    Lyu, Xiang
    Hu, Feifang
    STATISTICAL METHODS IN MEDICAL RESEARCH, 2019, 28 (01) : 184 - 195
  • [3] Optimizing Two-Way Partial AUC With an End-to-End Framework
    Yang, Zhiyong
    Xu, Qianqian
    Bao, Shilong
    He, Yuan
    Cao, Xiaochun
    Huang, Qingming
    IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2023, 45 (08) : 10228 - 10246
  • [4] The two-way Mundlak estimator
    Baltagi, Badi H.
    ECONOMETRIC REVIEWS, 2023, 42 (02) : 240 - 246
  • [5] The two-way Hausman and Taylor estimator
    Baltagi, Badi H.
    ECONOMICS LETTERS, 2023, 228
  • [6] When All We Need is a Piece of the Pie: A Generic Framework for Optimizing Two-way Partial AUC
    Yang, Zhiyong
    Xu, Qianqian
    Bao, Shilong
    He, Yuan
    Cao, Xiaochun
    Huang, Qingming
    INTERNATIONAL CONFERENCE ON MACHINE LEARNING, VOL 139, 2021, 139
  • [7] A Pretest Estimator for the Two-Way Error Component Model
    Baltagi, Badi H.
    Bresson, Georges
    Etienne, Jean-Michel
    ECONOMETRICS, 2024, 12 (02)
  • [8] Two-Way Interconnection with Partial Consumer Participation
    Aaron Schiff
    Networks and Spatial Economics, 2002, 2 (3) : 295 - 315
  • [9] Two-way pebble transducers for partial functions and their composition
    Joost Engelfriet
    Acta Informatica, 2015, 52 : 559 - 571
  • [10] Two-way pebble transducers for partial functions and their composition
    Engelfriet, Joost
    ACTA INFORMATICA, 2015, 52 (7-8) : 559 - 571