Hybridisation of Feature Selection and Classification Techniques in Credit Risk Assessment Modelling

被引：0

作者：

Sakri, Sapiah ^{[1
]}

Othman, Jaizah ^{[2
]}

Halid, Noreha ^{[3
]}

机构：

[1] Princess Nourah Bint Abdulrahman Univ, Coll Comp & Informat Sci, Riyadh 11671, Saudi Arabia

[2] Dublin City Univ, DCU Business Sch, Dublin 9, Ireland

[3] Princess Nourah Bint Abdulrahman Univ, Coll Business Adm, Riyadh 11671, Saudi Arabia

来源：

KNOWLEDGE INNOVATION THROUGH INTELLIGENT SOFTWARE METHODOLOGIES, TOOLS AND TECHNIQUES (SOMET_20) | 2020年 / 327卷

关键词：

Credit risk assessment; credit scoring prediction; ensemble classifiers; feature selection techniques; machine learning in credit risk assessment; single classifiers; GENETIC ALGORITHM; OPTIMIZATION; CLASSIFIERS; ENSEMBLES; SYSTEM;

D O I：

10.3233/FAIA200581

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

In recent years, the use of artificial intelligence techniques to manage credit risk has represented an improvement over conventional methods. Furthermore, small improvements to credit scoring systems and default forecasting can support huge profits. Accordingly, banks and financial institutions have a high interest in any changes. The literature shows that the use of feature selection techniques can reduce the dimensionality problems in most credit risk datasets, and, thus, improve the performance of the credit risk model. Many other works also indicated that various classification approaches would also affect the performance of the credit risk assessment modelling. In this research, based on the new proposed framework, we investigated the effect of various filter-based feature selection techniques with various classification approaches, namely, single and ensemble classifiers, on three credit datasets (German, Australian, and Japanese credit risk datasets) with the aim of improving the performance of the credit risk model. All single and ensemble classifier-based models were evaluated using four of the most used performance metrics for assessing financial stress models. From the comparison analysis between, with, and without applying the feature selection and across the three credit datasets, the Random-Forest + Information-Gain model achieved a better trade-off in improving the model's accuracy rate with the value of 96% for the Australian credit dataset. This model also obtained the lowest Type I error with the value of 4% for the German credit dataset, the lowest Type II error with the value of 2% for the German credit dataset and the highest value of G-mean of 95% for the Australian credit dataset. The results clearly indicate that the Random-Forest + Information-Gain model is an excellent predictor for the credit risk cases.

引用

页码：367 / 380

页数：14

共 50 条

[31] Imbalanced Data Classification Based on Feature Selection Techniques
Ksieniewicz, Pawel
Wozniak, Michal
INTELLIGENT DATA ENGINEERING AND AUTOMATED LEARNING (IDEAL 2018), PT II, 2018, 11315 : 296 - 303
[32] The effects of globalisation techniques on feature selection for text classification
Parlak, Bekir
Uysal, Alper Kursat
JOURNAL OF INFORMATION SCIENCE, 2021, 47 (06) : 727 - 739
[33] A Study On Feature Selection And Classification Techniques Of Indian Music
Kalapatapu, Prafulla
Goli, Srihita
Arthum, Prasanna
Malapati, Aruna
7TH INTERNATIONAL CONFERENCE ON EMERGING UBIQUITOUS SYSTEMS AND PERVASIVE NETWORKS (EUSPN 2016)/THE 6TH INTERNATIONAL CONFERENCE ON CURRENT AND FUTURE TRENDS OF INFORMATION AND COMMUNICATION TECHNOLOGIES IN HEALTHCARE (ICTH-2016), 2016, 98 : 125 - 131
[34] Improve Abstract Data with Feature Selection for Classification Techniques
Nuipian, Vatinee
Meesad, Phayung
Boonrawd, Pudsadee
MEMS, NANO AND SMART SYSTEMS, PTS 1-6, 2012, 403-408 : 3699 - +
[35] Feature selection for financial credit-risk evaluation decisions
Piramuthu, S
INFORMS JOURNAL ON COMPUTING, 1999, 11 (03) : 258 - 266
[36] Integrating data augmentation and hybrid feature selection for small sample credit risk assessment with high dimensionality
Zhang, Xiaoming
Yu, Lean
Yin, Hang
Lai, Kin Keung
Computers and Operations Research, 2022, 146
[37] Integrating data augmentation and hybrid feature selection for small sample credit risk assessment with high dimensionality
Zhang, Xiaoming
Yu, Lean
Yin, Hang
Lai, Kin Keung
COMPUTERS & OPERATIONS RESEARCH, 2022, 146
[38] A machine learning approach combining expert knowledge with genetic algorithms in feature selection for credit risk assessment
Lappas, Pantelis Z.
Yannacopoulos, Athanasios N.
APPLIED SOFT COMPUTING, 2021, 107
[39] The Bayesian Additive Classification Tree applied to credit risk modelling
Zhang, Junni L.
Haerdle, Wolfgang K.
COMPUTATIONAL STATISTICS & DATA ANALYSIS, 2010, 54 (05) : 1197 - 1205
[40] Adaptive Credit Card Fraud Detection Techniques Based on Feature Selection Method
Singh, Ajeet
Jain, Anurag
ADVANCES IN COMPUTER COMMUNICATION AND COMPUTATIONAL SCIENCES, IC4S 2018, 2019, 924 : 167 - 178

← 1 2 3 4 5 →