Author:
Behr Andreas,Weinblat Jurij
Abstract
Purpose
The purpose of this paper is to do a performance comparison of three different data mining techniques.
Design/methodology/approach
Logit model, decision tree and random forest are applied in this study on British, French, German, Italian, Portuguese and Spanish balance sheet data from 2006 to 2012, which covers 446,464 firms. Because of the strong imbalance with regard to the solvency status, classification trees and random forests are modified to adapt to this imbalance. All three model specifications are optimized extensively using resampling techniques, relying on the training sample only. Model performance is assessed, strictly, based on out-of-sample predictions.
Findings
Random forest is found to strongly outperform the classification tree and the logit model in almost all considered years and countries, according to the quality measure in this study.
Originality/value
Obtaining reliable estimates of default propensity scores is of immense importance for potential credit grantors, portfolio managers and regulatory authorities. As the overwhelming majority of firms are not listed on stock exchanges, annual balance sheets still provide the most important source of information. The obtained ranking of the three models according to their predictive performance is relatively robust, due to the consideration of several countries and a relatively long time period.
Reference51 articles.
1. Financial ratios, discriminant analysis and the prediction of corporate bankruptcy;The Journal of Finance,1968
2. Anders, U. and Szczesny, A. (1996), “Prognose von Insolvenzwahrscheinlichkeiten mit Hilfe logistischer neuronaler Netzwerke: Eine Untersuchung von kleinen und mittleren Unternehmen”, Technical Report, ZEW Discussion Papers.
3. Credit rating model development: an ordered analysis based on accounting data;Research in International Business and Finance,2016
4. Predicting probability of default of Indian corporate bonds: logistic and Z-score model approaches;The Journal of Risk Finance,2006
5. Do banks’ internal Basel risk estimates reflect risk?;Journal of Financial Stability,2014
Cited by
3 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献