Credit Risk Identification Using Machine Learning Models
سال انتشار: 1405
نوع سند: مقاله ژورنالی
زبان: فارسی
مشاهده: 55
فایل این مقاله در 8 صفحه با فرمت PDF قابل دریافت می باشد
- صدور گواهی نمایه سازی
- من نویسنده این مقاله هستم
استخراج به نرم افزارهای پژوهشی:
شناسه ملی سند علمی:
JR_IJDCS-9-1_001
تاریخ نمایه سازی: 14 مرداد 1405
چکیده مقاله:
Credit risk is one of the most fundamental challenges faced by financial institutions and banks in the loan approval process. Inaccurate assessment of borrowers' repayment capacity can lead to an increase in non-performing loans and impose significant financial losses on the banking system. Therefore, the application of advanced data analysis techniques and machine learning algorithms for accurate credit risk prediction has gained considerable importance.In this study, a credit risk assessment dataset was utilized, and a comprehensive data preprocessing procedure was performed. This process included handling and imputing missing values, identifying and removing outliers, normalizing numerical features, and encoding categorical variables to improve data quality for model training. Subsequently, three widely used machine learning algorithms, namely Extreme Gradient Boosting (XGBoost), Random Forest, and Support Vector Machine (SVM), were trained to predict loan repayment status.To evaluate the performance of the proposed models, several metrics were employed, including Accuracy, Precision, Recall, F۱-Score, and the Area Under the Receiver Operating Characteristic Curve (AUC-ROC), providing a comprehensive and multidimensional comparison of their predictive capabilities. The experimental results demonstrated that the XGBoost model outperformed the other models across most evaluation metrics and exhibited superior ability in correctly identifying both high-risk and low-risk customers. Based on these findings, it can be concluded that gradient boosting–based algorithms, particularly XGBoost, represent an efficient and reliable approach for credit risk prediction in financial institutions.
کلیدواژه ها:
Credit Risk ، ، Xgboost ، ، Random Forest ، ، Support Vector Machine (SVM) ، Data mining ، Loan Repayment Prediction
نویسندگان
میلاد قهاری بیدگلی
گروه مهندسی کامپیوتر، واحد اسلامشهر، دانشگاه آزاد اسلامی، تهران، ایران.
زهرا عباس نژاد
گروه مهندسی کامپیوتر، واحد علوم و تحقیقات، دانشگاه آزاد اسلامی، تهران، ایران.