Search | Korea Science

Semi-Supervised Learning to Predict Default Risk for P2P Lending (준지도학습 기반의 P2P 대출 부도 위험 예측에 대한 연구)

Kim, Hyun-jung
- Journal of Digital Convergence
- /
- v.20 no.4
- /
- pp.185-192
- /
- 2022
This study investigates the effect of the semi-supervised learning(SSL) method on predicting default risk of peer-to-peer(P2P) loans. Despite its proven performance, the supervised learning(SL) method requires labeled data, which may require a lot of effort and resources to collect. With the rapid growth of P2P platforms, the number of loans issued annually that have no clear final resolution is continuously increasing leading to abundance in unlabeled data. The research data of P2P loans used in this study were collected on the LendingClub platform. This is why an SSL model is needed to predict the default risk by using not only information from labeled loans(fully paid or defaulted) but also information from unlabeled loans. The results showed that in terms of default risk prediction and despite the use of a small number of labeled data, the SSL method achieved a much better default risk prediction performance than the SL method trained using a much larger set of labeled data.
https://doi.org/10.14400/JDC.2022.20.4.185 인용 PDF KSCI

Machine learning-based corporate default risk prediction model verification and policy recommendation: Focusing on improvement through stacking ensemble model (머신러닝 기반 기업부도위험 예측모델 검증 및 정책적 제언: 스태킹 앙상블 모델을 통한 개선을 중심으로)

Eom, Haneul;Kim, Jaeseong;Choi, Sangok
- Journal of Intelligence and Information Systems
- /
- v.26 no.2
- /
- pp.105-129
- /
- 2020
This study uses corporate data from 2012 to 2018 when K-IFRS was applied in earnest to predict default risks. The data used in the analysis totaled 10,545 rows, consisting of 160 columns including 38 in the statement of financial position, 26 in the statement of comprehensive income, 11 in the statement of cash flows, and 76 in the index of financial ratios. Unlike most previous prior studies used the default event as the basis for learning about default risk, this study calculated default risk using the market capitalization and stock price volatility of each company based on the Merton model. Through this, it was able to solve the problem of data imbalance due to the scarcity of default events, which had been pointed out as the limitation of the existing methodology, and the problem of reflecting the difference in default risk that exists within ordinary companies. Because learning was conducted only by using corporate information available to unlisted companies, default risks of unlisted companies without stock price information can be appropriately derived. Through this, it can provide stable default risk assessment services to unlisted companies that are difficult to determine proper default risk with traditional credit rating models such as small and medium-sized companies and startups. Although there has been an active study of predicting corporate default risks using machine learning recently, model bias issues exist because most studies are making predictions based on a single model. Stable and reliable valuation methodology is required for the calculation of default risk, given that the entity's default risk information is very widely utilized in the market and the sensitivity to the difference in default risk is high. Also, Strict standards are also required for methods of calculation. The credit rating method stipulated by the Financial Services Commission in the Financial Investment Regulations calls for the preparation of evaluation methods, including verification of the adequacy of evaluation methods, in consideration of past statistical data and experiences on credit ratings and changes in future market conditions. This study allowed the reduction of individual models' bias by utilizing stacking ensemble techniques that synthesize various machine learning models. This allows us to capture complex nonlinear relationships between default risk and various corporate information and maximize the advantages of machine learning-based default risk prediction models that take less time to calculate. To calculate forecasts by sub model to be used as input data for the Stacking Ensemble model, training data were divided into seven pieces, and sub-models were trained in a divided set to produce forecasts. To compare the predictive power of the Stacking Ensemble model, Random Forest, MLP, and CNN models were trained with full training data, then the predictive power of each model was verified on the test set. The analysis showed that the Stacking Ensemble model exceeded the predictive power of the Random Forest model, which had the best performance on a single model. Next, to check for statistically significant differences between the Stacking Ensemble model and the forecasts for each individual model, the Pair between the Stacking Ensemble model and each individual model was constructed. Because the results of the Shapiro-wilk normality test also showed that all Pair did not follow normality, Using the nonparametric method wilcoxon rank sum test, we checked whether the two model forecasts that make up the Pair showed statistically significant differences. The analysis showed that the forecasts of the Staging Ensemble model showed statistically significant differences from those of the MLP model and CNN model. In addition, this study can provide a methodology that allows existing credit rating agencies to apply machine learning-based bankruptcy risk prediction methodologies, given that traditional credit rating models can also be reflected as sub-models to calculate the final default probability. Also, the Stacking Ensemble techniques proposed in this study can help design to meet the requirements of the Financial Investment Business Regulations through the combination of various sub-models. We hope that this research will be used as a resource to increase practical use by overcoming and improving the limitations of existing machine learning-based models.
https://doi.org/10.13088/jiis.2020.26.2.105 인용 PDF KSCI

A Comparative Analysis for the knowledge of Data Mining Techniques with Experties (Data Mining 기법들과 전문가들로부터 추출된 지식에 관한 실증적 비교 연구)

김광용;손광기;홍온선
- Journal of Intelligence and Information Systems
- /
- v.4 no.1
- /
- pp.41-58
- /
- 1998
본 연구는 여러 가지 Data Mining 기법들로부터 도출된 지식과 AHP를 이용하여 도출된 전문가의 지식을 사용된 정보의 특성에 따라 조사하고, 이러한 각각의 지식들을 중심으로 부도예측 모형을 설계한 후, 각 모형의 특성 및 부도예측력에 대한 실증적 비교연구에 그 목적을 두고 있다. 사용된 Data Mining 기법들은 통계적 다중판별분석 모형, ID3 모형, 인공신경망 모형이며, 전문가 지식의 추출은 AHP를 사용하여 45명의 전문가로부터 부도와 관련하여 인터뷰 및 설문조사를 실시하였다. 특히 부도예측에 사용된 변수의 특성을 정량적 재무정보와 정성적 비재무정보로 나누어서 각 모형의 특성을 비교연구하였다. 연구결과 부도예측시 정성적정보의 중요성을 확인하였으며, 전문가의 지식을 기반으로한 AHP 모형이 위험예측모형으로 사용될 수 있음을 실증적으로 보여주었다.
PDF

SOHO Bankruptcy Prediction Using Modified Bagging Predictors (Modified Bagging Predictors를 이용한 SOHO 부도 예측)

Kim Seung-Hyeok;Kim Jong-U
- Proceedings of the Korea Inteligent Information System Society Conference
- /
- 2006.06a
- /
- pp.176-182
- /
- 2006
본 연구에서는 기존 Bagging Predictors에 수정을 가한 Modified Bagging Predictors를 이용하여 SOHO 에 대한 부도예측 모델을 제시한다. 대기업 및 중소기업에 대한 기압부도예측 모델에 대한 많은 선행 연구가 있어왔지만 SOHO 만의 기업부도예측 모델에 관한 연구는 미비한 상태이다. 금융기관들의 대출심사시 대기업 및 중소기업과는 달리 SOHO에 대한 대출심사는 이직은 체계화되지 못한 채 신용정보점수 등의 단편적인 요소를 사용하고 있는 것에 현실이고 이에 따라 잘못된 대출로 안한 금융기관의 부실화를 초래할 위험성이 크다. 본 연구에서는 실제 국내은행의 SOHO 데이터 집합이 사용되었다. 먼저 기업부도 예측 모델에서 우수하다고 연구되어진 인공신경망과 의사결정나무 추론 기법을 적용하여 보았지만 만족할 만한 성과를 이쓸어내지 못하여, 기존 기업부도예측 모델연구에서 적용이 미비하였던 Bagging Predictors와 이를 개선한 Modified Bagging Predictors를 제시하고 이를 적용하여 보았다. 연구결과,; SOHO 부도예측에 있어서 본 연구에서 제시한 Modified Bagging Predictors 가 인공신경망과 Bagging Predictors등의 기존 기법에 비해서 성과가 향상됨을 알 수 있었다. 제시된 Modified Bagging Predictors의 유용성을 확인하기 위해서 추가적으로 대수의 공개 데이터 집합을 활용하여 성능을 비교한 결과 Modified Bagging Predictors 가 기존의 Bagging Predictors 에 비해 일관적으로 성과가 향상됨을 알 수 있었다.
PDF

Option-type Default Forecasting Model of a Firm Incorporating Debt Structure, and Credit Risk (기업의 부채구조를 고려한 옵션형 기업부도예측모형과 신용리스크)

Won, Chae-Hwan;Choi, Jae-Gon
- The Korean Journal of Financial Management
- /
- v.23 no.2
- /
- pp.209-237
- /
- 2006
Since previous default forecasting models for the firms evaluate the probability of default based upon the accounting data from book values, they cannot reflect the changes in markets sensitively and they seem to lack theoretical background. The market-information based models, however, not only make use of market data for the default prediction, but also have strong theoretical background like Black-Scholes (1973) option theory. So, many firms recently use such market based model as KMV to forecast their default probabilities and to manage their credit risks. Korean firms also widely use the KMV model in which default point is defined by liquid debt plus 50% of fixed debt. Since the debt structures between Korean and American firms are significantly different, Korean firms should carefully use KMV model. In this study, we empirically investigate the importance of debt structure. In particular, we find the following facts: First, in Korea, fixed debts are more important than liquid debts in accurate prediction of default. Second, the percentage of fixed debt must be less than 20% when default point is calculated for Korean firms, which is different from the KMV. These facts give Korean firms some valuable implication about default forecasting and management of credit risk.
PDF

Using Business Failure Probability Map (BFPM) for Corporate Credit Rating (다중 부실예측모형을 이용한 통합 신용등급화 방법)

신택수;홍태호
- Proceedings of the Korean Operations and Management Science Society Conference
- /
- 2003.05a
- /
- pp.835-842
- /
- 2003
현행 기업신용평가모형에 관한 연구는 크게 부실예측모형 및 채권등급 평가모형으로 구분된다. 이러한 신응평가모형에 관한 연구는 단순히 부실여부 또는 이미 전문가 집단에 의해 사전에 정의된 등급체계만을 예측하는 데 초점을 맞추고 있었다. 그러나. 대부분의 금융기관에서 사용하는 신응평가모형은 기업의 부실여부만을 예측하거나 기존의 채권등급을 예측하기 위만 목적보다는 기업의 고유 신응위험을 평가하여 이에 적합한 신용등급을 부여함으로써, 효율적인 대출업무를 수행하기 위해 활용되고 있다. 본 연구에서는 기존의 부실예측모형들을 대상으로 다중 부실확률모형 (Business Failure Probability Map; BFPM) 접근방법을 이용한 신응등급화 방법을 제안하고자 한다. 본 연구에서 제시된 다중 부실확률모형은 신경망모형과 로짓모형을 통합하여 부도율, 점유율을 고려한 다단계 신용등급을 예측할 수 있게 해준다. 다중 부도확률지도 접근방법을 이용하여 각 금융기관에서 정의하는 수준의 신용리스크를 효과적으로 추정하고, 이를 기준으로 보다 객관적인 다단계 신용등급을 산출하는 새로운 신응등급화 방법을 제시 하고자 한다.
PDF

Predicting Default of Construction Companies Using Bayesian Probabilistic Approach (베이지안 확률적 접근법을 이용한 건설업체 부도 예측에 관한 연구)

Hong, Sungmoon;Hwang, Jaeyeon;Kwon, Taewhan;Kim, Juhyung;Kim, Jaejun
- Korean Journal of Construction Engineering and Management
- /
- v.17 no.5
- /
- pp.13-21
- /
- 2016
Insolvency of construction companies that play the role of main contractors can lead to clients' losses due to non-fulfillment of construction contracts, and it can have negative effects on the financial soundness of construction companies and suppliers. The construction industry has the cash flow financial characteristic of receiving a project and getting payment based on the progress of the construction. As such, insolvency during project progress can lead to financial losses, which is why the prediction of construction companies is so important. The prediction of insolvency of Korean construction companies are often made through the KMV model from the KMV (Kealhofer McQuown and Vasicek) Company developed in the U.S. during the early 90s, but this model is insufficient in predicting construction companies because it was developed based on credit risk assessment of general companies and banks. In addition, the predictive performance of KMV value's insolvency probability is continuously being questioned due to lack of number of analyzed companies and data. Therefore, in order to resolve such issues, the Bayesian Probabilistic Approach is to be combined with the existing insolvency predictive probability model. This is because if the Prior Probability of Bayesian statistics can be appropriately predicted, reliable Posterior Probability can be predicted through ensured conditionality on the evidence despite the lack of data. Thus, this study is to measure the Expected Default Frequency (EDF) by utilizing the Bayesian Probabilistic Approach with the existing insolvency predictive probability model and predict the accuracy by comparing the result with the EDF of the existing model.
https://doi.org/10.6106/KJCEM.2016.17.5.013 인용 PDF KSCI

SOHO Bankruptcy Prediction Using Modified Bagging Predictors (Modified Bagging Predictors를 이용한 SOHO 부도 예측)

Kim, Seung-Hyuk;Kim, Jong-Woo
- Journal of Intelligence and Information Systems
- /
- v.13 no.2
- /
- pp.15-26
- /
- 2007
In this study, a SOHO (Small Office Home Office) bankruptcy prediction model is proposed using Modified Bagging Predictors which is modification of traditional Bagging Predictors. There have been several studies on bankruptcy prediction for large and middle size companies. However, little studies have been done for SOHOs. In commercial banks, loan approval processes for SOHOs are usually less structured than those for large and middle size companies, and largely depend on partial information such as credit scores. In this study, we use a real SOHO loan approval data set of a Korean bank. First, decision tree induction techniques and artificial neural networks are applied to the data set, and the results are not satisfactory. Bagging Predictors which has been not previously applied for bankruptcy prediction and Modified Bagging Predictors which is proposed in this paper are applied to the data set. The experimental results show that Modified Bagging Predictors provides better performance than decision tree inductions techniques, artificial neural networks, and Bagging Predictors.
PDF

An Empirical Study on the Risk Index of Korean Securities Industry (우리나라 증권산업의 위험지수 작성에 관한 실증연구)

Chang, Kook-Hyun
- The Korean Journal of Financial Management
- /
- v.25 no.3
- /
- pp.131-153
- /
- 2008
This paper calculates the Risk Index of Korean securities industry that summarizes the information contained in seventeen financial indicators that represent risk categories such as capital adequacy(C), asset quality(A), earnings(E), and liquidity(L) by using the NBER statistical methodology. For the validation of Risk Index, expected default frequency has been used, and the result has been proved to be positive. According to the compiled Risk Index, the level of risks of Korean securities industry has been decreasing from the second quarter of 2003 to the first quarter of 2006 by 22 percent. But the risk has been increasing during the periods from the first quarter of 2002 to the first quarter of 2003 and from the first quarter of 2006 to the last quarter of 2006.
PDF

A Study on the Sustainability of New SMEs through the Analysis of Altman Z-Score: Focusing on New and Renewable Energy Industry in Korea (알트만 Z-스코어를 이용한 신생 중소기업의 지속가능성 분석: 신재생에너지산업을 중심으로)

Oh, Nak-Kyo;Yoon, Sung-Soo;Park, Won-Koo
- Journal of Technology Innovation
- /
- v.22 no.2
- /
- pp.185-220
- /
- 2014
The purpose of this study is to get a whole picture of financial conditions of the new and renewable energy sector which have been growing rapidly and predict bankruptcy risk quantitatively. There have been many researches on the methodologies for company failure prediction, such as financial ratios as predictors of failure, analysis of corporate governance, risk factors and survival analysis, and others. The research method for this study is Altman Z-score which has been widely used in the world. Data Set was composed of 121 companies with financial statements from KIS-Value. Covering period for the analysis of the data set is from the year 2006 to 2011. As a result of this study, we found that 38 percent of the data set belongs to "Distress" Zone (on alert) while 38% (on watch), summed into 76%, whose level could be interpreted to doubt about the sustainability. The average of the SMEs in wind energy sector was worse than that of SMEs in solar energy sector. And the average of the SMEs in the "Distress" Zone (on alert) was worse than that of the companies of large group in the "Distress" Zone (on alert). In conclusion, Altman Z-score was well proved to be effective for New & Renewable Energy Industry in Korea as a result of this study. The importance of this study lies on the result to demonstrate empirically that the majority of solar and wind enterprises are facing the risk of bankruptcy. And it is also meaningful to have studied the relationship between SMEs and large companies in addition to advancing research on new start-up companies.
https://doi.org/10.14383/SIME.2014.22.2.185 인용 PDF

Search Result 16, Processing Time 0.026 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)