Explainability of Machine Learning Models in Credit Risk Management

dc.contributor.advisorMukhodobwane, R. M.
dc.contributor.advisorMphephu, N.
dc.contributor.advisorNetshikweta, R.
dc.contributor.authorDzhivhuho, Asikundwi Praise the Lord
dc.date2026
dc.date.accessioned2026-06-17T21:24:56Z
dc.date.available2026-06-17T21:24:56Z
dc.date.issued2026-05-19
dc.descriptionM.Sc. in e-Science
dc.descriptionDepartment of Mathematical and Computational Sciences
dc.description.abstractThe effective management of credit risk is a critical challenge for financial institutions, with accurate assessment of loan default risk playing a central role in maintaining financial stability. Machine Learning (ML) techniques have become increasingly prevalent in credit risk assessment due to their ability to capture complex patterns in borrower behavior and improve predictive accuracy. However, the lack of interpretability of many advanced ML models, such as Random Forest, XGBoost, and Neural Networks, raises concerns regarding transparency, fairness, and accountability in decision-making, particularly in high-stakes environments where regulatory compliance and ethical considerations are paramount. This study seeks to bridge the gap between predictive accuracy and interpretability by applying two post-hoc, model-agnostic explainability techniques Local Interpretable Model-Agnostic Explanations (LIME) and SHapley Additive exPlanations (SHAP) to evaluate five commonly used ML models Logistic Regression, Multivariate Adaptive Regression Splines (MARS), Neural Networks, Random Forest, and XGBoost. Using an open-access Kaggle dataset, the study examines both the predictive performance and the interpretability of these models, with a particular focus on the trade-offs between high accuracy and model transparency. The results highlight a clear trade-off while ensemble models like XGBoost and Random Forests exhibit superior accuracy, particularly in predicting low-risk borrowers, they struggle with detecting high risk applicants and lack the interpretability required for transparent decision-making. Simpler models, such as Logistic Regression, offer greater transparency and are more effective in identifying high-risk cases but sacrifice predictive accuracy. Neural Networks strike a balance, providing better accuracy than linear models while maintaining moderate sensitivity to high-risk applicants. By leveraging SHAP and LIME, this research enhances model transparency, offering both global insights into risk factors and local instance-level explanations for individual predictions, which can aid stakeholders such as financial institutions, regulators, and applicants in making more informed, fair, and accountable credit decision
dc.format.extent1 online resource (x, 99 leaves)
dc.identifier.apacitationDzhivhuho, A. P. t. L. (2026). <i>Explainability of Machine Learning Models in Credit Risk Management</i>. (). . Retrieved from en_ZA
dc.identifier.chicagocitationDzhivhuho, Asikundwi Praise the Lord. <i>"Explainability of Machine Learning Models in Credit Risk Management."</i> ., , 2026. en_ZA
dc.identifier.citationDzhivhuho, A.P.t.L. 2026. Explainability of Machine Learning Models in Credit Risk Management. . . en_ZA
dc.identifier.ris TY - Dissertation AU - Dzhivhuho, Asikundwi Praise the Lord AB - The effective management of credit risk is a critical challenge for financial institutions, with accurate assessment of loan default risk playing a central role in maintaining financial stability. Machine Learning (ML) techniques have become increasingly prevalent in credit risk assessment due to their ability to capture complex patterns in borrower behavior and improve predictive accuracy. However, the lack of interpretability of many advanced ML models, such as Random Forest, XGBoost, and Neural Networks, raises concerns regarding transparency, fairness, and accountability in decision-making, particularly in high-stakes environments where regulatory compliance and ethical considerations are paramount. This study seeks to bridge the gap between predictive accuracy and interpretability by applying two post-hoc, model-agnostic explainability techniques Local Interpretable Model-Agnostic Explanations (LIME) and SHapley Additive exPlanations (SHAP) to evaluate five commonly used ML models Logistic Regression, Multivariate Adaptive Regression Splines (MARS), Neural Networks, Random Forest, and XGBoost. Using an open-access Kaggle dataset, the study examines both the predictive performance and the interpretability of these models, with a particular focus on the trade-offs between high accuracy and model transparency. The results highlight a clear trade-off while ensemble models like XGBoost and Random Forests exhibit superior accuracy, particularly in predicting low-risk borrowers, they struggle with detecting high risk applicants and lack the interpretability required for transparent decision-making. Simpler models, such as Logistic Regression, offer greater transparency and are more effective in identifying high-risk cases but sacrifice predictive accuracy. Neural Networks strike a balance, providing better accuracy than linear models while maintaining moderate sensitivity to high-risk applicants. By leveraging SHAP and LIME, this research enhances model transparency, offering both global insights into risk factors and local instance-level explanations for individual predictions, which can aid stakeholders such as financial institutions, regulators, and applicants in making more informed, fair, and accountable credit decision DA - 2026-05-19 DB - ResearchSpace DP - Univen KW - Machine Learning KW - Credit Risk KW - Interpretability KW - SHAP KW - LIME KW - Logistic Regression KW - MARS KW - Neural Networks KW - Random Forest KW - XGboost LK - https://univendspace.univen.ac.za PY - 2026 T1 - Explainability of Machine Learning Models in Credit Risk Management TI - Explainability of Machine Learning Models in Credit Risk Management UR - ER - en_ZA
dc.identifier.urihttps://univendspace.univen.ac.za/handle/11602/3211
dc.identifier.vancouvercitationDzhivhuho APtL. Explainability of Machine Learning Models in Credit Risk Management. []. , 2026 [cited yyyy month dd]. Available from: en_ZA
dc.language.isoen
dc.relation.requiresPDF
dc.rightsUniversity of Venda
dc.subjectMachine Learning
dc.subjectUCTDen_ZA
dc.subjectInterpretability
dc.subjectSHAP
dc.subjectLIME
dc.subjectLogistic Regression
dc.subjectMARS
dc.subjectNeural Networks
dc.subjectRandom Forest
dc.subjectXGboost
dc.titleExplainability of Machine Learning Models in Credit Risk Management
dc.typeDissertation

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Dissertation - Dzhivhuho, a. p.-.pdf
Size:
2.36 MB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed upon to submission
Description: