• Title/Summary/Keyword: Logistic regression analysis

Search Result 4,143, Processing Time 0.043 seconds

Semiparametric kernel logistic regression with longitudinal data

  • Shim, Joo-Yong;Seok, Kyung-Ha
    • Journal of the Korean Data and Information Science Society
    • /
    • v.23 no.2
    • /
    • pp.385-392
    • /
    • 2012
  • Logistic regression is a well known binary classification method in the field of statistical learning. Mixed-effect regression models are widely used for the analysis of correlated data such as those found in longitudinal studies. We consider kernel extensions with semiparametric fixed effects and parametric random effects for the logistic regression. The estimation is performed through the penalized likelihood method based on kernel trick, and our focus is on the efficient computation and the effective hyperparameter selection. For the selection of optimal hyperparameters, cross-validation techniques are employed. Numerical results are then presented to indicate the performance of the proposed procedure.

FACTORS AFFECTING PATIENTS' DECISION-MAKING FOR DENTAL PROSTHETIC TREATMENT

  • Jung, Hyo-Kyung;Kim, Han-Gon
    • The Journal of Korean Academy of Prosthodontics
    • /
    • v.46 no.6
    • /
    • pp.610-619
    • /
    • 2008
  • STATEMENT OF PROBLEM: Factors affecting patients' decision-making for dental prosthetic treatment should be examined in terms of understanding improving patients' oral health. PURPOSE: The main purpose of this dissertation was to investigate patients' dental prosthetic treatment and factors affecting patients' decision-making for dental prosthesis treatment in Deagu and Gyungbook areas. MATERIAL AND METHODS: This study was based on the preliminary survey of dental patients conducted from July 1 to August 31 in 2006. A total of 700 questionnaires had been distributed and 640 were collected. 629 questionnaires were used for the statistical analysis. Descriptive and inferential statistics, such as frequencies, cross tabulation analysis, correlation analysis, logistic regression analysis, and multiple regression analysis were introduced. In the multiple regression analysis and logistic regression analysis, twenty-two independent variables were employed to explore the factors which have impacts on decision-making and satisfaction. RESULTS: The results of this dissertation are as follows: Logistic regression analysis turned out that monthly income, age, degree of expectation, marital status, and employer-insured policy of national insurance statistically increased the odds of decision-making of dental prosthesis treatment. But educational attainment decreased the odds ratio of the decision-making of dental prosthesis treatment. However, the rest independent variables do not have statistically significant impacts on the decision-making of dental prosthesis treatment CONCLUSION: Among independent variables, marital status had the most significant influence on the decision making of dental prosthesis treatment. Finally, suggestions for the future study and policy implications to improve satisfaction of the patients' dental prosthetic treatment were discussed.

A Study on the Development of Product Planning Prediction Model Using Logistic Regression Algorithm (로지스틱 회귀 알고리즘을 활용한 상품 기획 예측 모형 개발에 관한 연구)

  • Ahn, Yeong-Hwil;Park, Koo-Rack;Kim, Dong-Hyun;Kim, Do-Yeon
    • Journal of the Korea Convergence Society
    • /
    • v.12 no.9
    • /
    • pp.39-47
    • /
    • 2021
  • This study was conducted to propose a product planning prediction model using logistic regression algorithm to predict seasonal factors and rapidly changing product trends. First, we collected unstructured data of consumers in portal sites and online markets using web crawling, and analyzed meaningful information about products through preprocessing for transformation of standardized data. The datasets of 11,200 were analyzed by Logistic Regression to analyze consumer satisfaction, frequency analysis, and advantages and disadvantages of products. The result of analysis showed that the satisfaction of consumers was 92% and the defective issues of products were confirmed through frequency analysis. The results of analysis on the use satisfaction, system efficiency, and system effectiveness items of the developed product planning prediction program showed that the satisfaction was high. Defective issues are very meaningful data in that they provide information necessary for quickly recognizing the current problem of products and establishing improvement strategies.

A Study on Factors Affecting the Use of Ambulatory Physician Services (의사방문수 결정요인 분석)

  • 박현애;송건용
    • Health Policy and Management
    • /
    • v.4 no.2
    • /
    • pp.58-76
    • /
    • 1994
  • In order to study factors affecting the use of the ambulatory physician services. Andersen's model for health utilization was modified by adding the health behavior component and examined with three different approaches. Three different approaches were the multiople regression model, logistic regression model, and LISREL model. For multiple regression, dependent variable was reported illness-related visits to a physician during past one year and independent variables are variaous variables measuring predisposing factor, enabling factor, need factor and health behavior. For the logistic regression, dependent variable was visit or no-visit to a physician during past one year and independent variables were same as the multiple regression analysis. For the LISREL, five endogenous variables of health utiliztion, predisposing factor, enabling factor, need factor, and health behavior and 20 exogeneous variables which measures five endogenous variables were used. According to the multiple regression analysis, chronic illness, health status, perceived health status of the need factor; residence, sex, age, marital status, education of the predisposing factor ; health insurance, usual source for medical care of enabling factor were the siginificant exploratory variables for the health utilization. Out of the logistic regression analysis, health status, chronic illness, residence, marital status, education, drinking, use of health aid were found to be significant exploratory variables. From LISREL, need factor affect utilization most following by predisposing factor, enabling factor and health behavior. For LISREL model, age, education, and residence for predisposing factor; health status, chronic illess, and perceived health status for need factor; medical insurance for enabling factor; and doing any kind of health behavior for the health behavior were found as the significant observed variables for each theoretical variables.

  • PDF

Variable Selection for Logistic Regression Model Using Adjusted Coefficients of Determination (수정 결정계수를 사용한 로지스틱 회귀모형에서의 변수선택법)

  • Hong C. S.;Ham J. H.;Kim H. I.
    • The Korean Journal of Applied Statistics
    • /
    • v.18 no.2
    • /
    • pp.435-443
    • /
    • 2005
  • Coefficients of determination in logistic regression analysis are defined as various statistics, and their values are relatively smaller than those for linear regression model. These coefficients of determination are not generally used to evaluate and diagnose logistic regression model. Liao and McGee (2003) proposed two adjusted coefficients of determination which are robust at the addition of inappropriate predictors and the variation of sample size. In this work, these adjusted coefficients of determination are applied to variable selection method for logistic regression model and compared with results of other methods such as the forward selection, backward elimination, stepwise selection, and AIC statistic.

Prediction on Busan's Gross Product and Employment of Major Industry with Logistic Regression and Machine Learning Model (로지스틱 회귀모형과 머신러닝 모형을 활용한 주요산업의 부산 지역총생산 및 고용 효과 예측)

  • Chae-Deug Yi
    • Korea Trade Review
    • /
    • v.47 no.2
    • /
    • pp.69-88
    • /
    • 2022
  • This paper aims to predict Busan's regional product and employment using the logistic regression models and machine learning models. The following are the main findings of the empirical analysis. First, the OLS regression model shows that the main industries such as electricity and electronics, machine and transport, and finance and insurance affect the Busan's income positively. Second, the binomial logistic regression models show that the Busan's strategic industries such as the future transport machinery, life-care, and smart marine industries contribute on the Busan's income in large order. Third, the multinomial logistic regression models show that the Korea's main industries such as the precise machinery, transport equipment, and machinery influence the Busan's economy positively. And Korea's exports and the depreciation can affect Busan's economy more positively at the higher employment level. Fourth, the voting ensemble model show the higher predictive power than artificial neural network model and support vector machine models. Furthermore, the gradient boosting model and the random forest show the higher predictive power than the voting model in large order.

Developing a Combined Forecasting Model on Hospital Closure (병원도산의 예측모형 개발연구)

  • 정기택;이훈영
    • Health Policy and Management
    • /
    • v.10 no.2
    • /
    • pp.1-21
    • /
    • 2000
  • This study reviewde various parametic and nonparametic method for forexasting hospital closures in Korea. We compared multivariate discriminant analysis, multivartiate logistic regression, classfication and regression tree, and neural network method based on hit ratio of each model for forecasting hospital closure. Like other studies in the literture, neural metwork analysis showed highest average hit ratio. For policy and business purposes, we combined the four analytical method and constructed a foreasting model that can be easily used to predict the probabolity of hospital closure given financial information of a hospital.

  • PDF

Analysis of Landslide Hazard Area using Logistic Regression Analysis and AHP (Analytical Hierarchy Process) Approach (로지스틱 회귀분석 및 AHP 기법을 이용한 산사태 위험지역 분석)

  • Lee, Yong-jun;Park, Geun-Ae;Kim, Seong-Joon
    • KSCE Journal of Civil and Environmental Engineering Research
    • /
    • v.26 no.5D
    • /
    • pp.861-867
    • /
    • 2006
  • The objective of this study is to analyze the landslide hazard areas by combining LRA (Lgistic Regression Analysis) and AHP (Analytic Hierarchy Program) methods with Remote Sensing and GIS data in Anseong-si. In order to classify landslide hazard areas of seven levels, six topographic factors (slope, aspect, elevation, soil drain, soil depth, and land use) were used as input factors of LRA and AHP methods. As results, high-risk areas for landslide (1 and 2 levels) by LRA and AHP of its own were classified as 46.1% and 48.7%, respectively. A new method by applying weighting factors to the results of LRA and AHP was suggested. High-risk areas for landslide (1 and 2 levels) form the new method was classified as 58.9%.

Comparative Analysis of Predictors of Depression for Residents in a Metropolitan City using Logistic Regression and Decision Making Tree (로지스틱 회귀분석과 의사결정나무 분석을 이용한 일 대도시 주민의 우울 예측요인 비교 연구)

  • Kim, Soo-Jin;Kim, Bo-Young
    • The Journal of the Korea Contents Association
    • /
    • v.13 no.12
    • /
    • pp.829-839
    • /
    • 2013
  • This study is a descriptive research study with the purpose of predicting and comparing factors of depression affecting residents in a metropolitan city by using logistic regression analysis and decision-making tree analysis. The subjects for the study were 462 residents ($20{\leq}aged{\angle}65$) in a metropolitan city. This study collected data between October 7, 2011 and October 21, 2011 and analyzed them with frequency analysis, percentage, the mean and standard deviation, ${\chi}^2$-test, t-test, logistic regression analysis, roc curve, and a decision-making tree by using SPSS 18.0 program. The common predicting variables of depression in community residents were social dysfunction, perceived physical symptom, and family support. The specialty and sensitivity of logistic regression explained 93.8% and 42.5%. The receiver operating characteristic (roc) curve was used to determine an optimal model. The AUC (area under the curve) was .84. Roc curve was found to be statistically significant (p=<.001). The specialty and sensitivity of decision-making tree analysis were 98.3% and 20.8% respectively. As for the whole classification accuracy, the logistic regression explained 82.0% and the decision making tree analysis explained 80.5%. From the results of this study, it is believed that the sensitivity, the classification accuracy, and the logistics regression analysis as shown in a higher degree may be useful materials to establish a depression prediction model for the community residents.

A Study on Life Cycle analysis and prediction of Contents Service in the Wireless Internet (로지스틱 회귀 모형을 이용한 무선인터넷 콘텐츠 서비스의 life cycle 분석 및 예측)

  • Park, Ji-Hong;Jeon, Joon-Hyeon
    • Proceedings of the IEEK Conference
    • /
    • 2005.11a
    • /
    • pp.1161-1164
    • /
    • 2005
  • In this paper, we proposed the technique to estimate the life cycle of Internet content services based on the logistic regression model. In this paper, to define parameters of Internet contents estimating life cycle by logistic regression model, we used market size, traffic amount, page view and session-visit number as the parameters of Internet contents estimating life cycle by logistic regression model. In this paper, to compare the performance of our proposed scheme, we estimated life cycle for the download services of bell sound & character contents in mobile network. As a result, using our proposed logistic regression, we were able to estimate exactly the life cycle of the download services of bell sound & character contents.

  • PDF