A machine learning-based approach to prognostic analysis of thoracic transplantations

Authors:
Dursun Delen;Asil Oztekin;Zhenyu (James) Kong
Affiliations:
Spears School of Business, Oklahoma State University, T-NCB 378, 700 North Greenwood Avenue, Tulsa, OK, 74106, USA;School of Industrial Engineering and Management, Oklahoma State University, 322 Engineering North, Stillwater, OK 74078, USA and Department of Industrial Engineering, Gediz University, 35230 Canka ...;School of Industrial Engineering and Management, Oklahoma State University, 322 Engineering North, Stillwater, OK 74078, USA
Venue:
Artificial Intelligence in Medicine
Year:
2010

Citing 15
Cited 5

Applied multivariate statistical analysis

Applied multivariate statistical analysis
Multivariate data analysis (4th ed.): with readings

Multivariate data analysis (4th ed.): with readings
Clustering techniques

Future Generation Computer Systems - Special double issue on data mining
An introduction to support Vector Machines: and other kernel-based learning methods

An introduction to support Vector Machines: and other kernel-based learning methods
A robust and scalable clustering algorithm for mixed type attributes in large database environment

Proceedings of the seventh ACM SIGKDD international conference on Knowledge discovery and data mining
Organ transplantation policies: an update on a successful simulation project: the Unos Liver Allocation Model

Proceedings of the 32nd conference on Winter simulation
Neural Networks: A Comprehensive Foundation

Neural Networks: A Comprehensive Foundation
Machine Learning

Machine Learning
Modeling medical prognosis: survival analysis techniques

Computers and Biomedical Research
Logistic regression and artificial neural network classification models: a methodology review

Journal of Biomedical Informatics
Sensitivity Analysis in Practice: A Guide to Assessing Scientific Models

Sensitivity Analysis in Practice: A Guide to Assessing Scientific Models
Single and multiple time-point prediction models in kidney transplant outcomes

Journal of Biomedical Informatics
A study of cross-validation and bootstrap for accuracy estimation and model selection

IJCAI'95 Proceedings of the 14th international joint conference on Artificial intelligence - Volume 2
Advanced Data Mining Techniques

Advanced Data Mining Techniques
Genetic K-means algorithm

IEEE Transactions on Systems, Man, and Cybernetics, Part B: Cybernetics

Development of a structural equation modeling-based decision tree methodology for the analysis of lung transplantations

Decision Support Systems
An analytic approach to better understanding and management of coronary surgeries

Decision Support Systems
Predicting patient survival after liver transplantation using evolutionary multi-objective artificial neural networks

Artificial Intelligence in Medicine
Predicting re-hospitalisations using intelligent systems: an exploratory study

International Journal of Business Information Systems
Review: Knowledge discovery in medicine: Current issue and future trend

Expert Systems with Applications: An International Journal

Quantified Score

Hi-index	0.00

Visualization

Abstract

Objective: The prediction of survival time after organ transplantations and prognosis analysis of different risk groups of transplant patients are not only clinically important but also technically challenging. The current studies, which are mostly linear modeling-based statistical analyses, have focused on small sets of disparate predictive factors where many potentially important variables are neglected in their analyses. Data mining methods, such as machine learning-based approaches, are capable of providing an effective way of overcoming these limitations by utilizing sufficiently large data sets with many predictive factors to identify not only linear associations but also highly complex, non-linear relationships. Therefore, this study is aimed at exploring risk groups of thoracic recipients through machine learning-based methods. Methods and material: A large, feature-rich, nation-wide thoracic transplantation dataset (obtained from the United Network for Organ Sharing-UNOS) is used to develop predictive models for the survival time estimation. The predictive factors that are most relevant to the survival time identified via, (1) conducting sensitivity analysis on models developed by the machine learning methods, (2) extraction of variables from the published literature, and (3) eliciting variables from the medical experts and other domain specific knowledge bases. A unified set of predictors is then used to develop a Cox regression model and the related prognosis indices. A comparison of clustering algorithm-based and conventional risk grouping techniques is conducted based on the outcome of the Cox regression model in order to identify optimal number of risk groups of thoracic recipients. Finally, the Kaplan-Meier survival analysis is performed to validate the discrimination among the identified various risk groups. Results: The machine learning models performed very effectively in predicting the survival time: the support vector machine model with a radial basis Kernel function produced the best fit with an R^2 value of 0.879, the artificial neural network (multilayer perceptron-MLP-model) came the second with an R^2 value of 0.847, and the M5 algorithm-based regression tree model came last with an R^2 value of 0.785. Following the proposed method, a consolidated set of predictive variables are determined and used to build the Cox survival model. Using the prognosis indices revealed by the Cox survival model along with a k-means clustering algorithm, an optimal number of ''three'' risk groups is identified. The significance of differences among these risk groups are also validated using the Kaplan-Meier survival analysis. Conclusions: This study demonstrated that the integrated machine learning method to select the predictor variables is more effective in developing the Cox survival models than the traditional methods commonly found in the literature. The significant distinction among the risk groups of thoracic patients also validates the effectiveness of the methodology proposed herein. We anticipate that this study (and other AI based analytic studies like this one) will lead to more effective analyses of thoracic transplant procedures to better understand the prognosis of thoracic organ recipients. It would potentially lead to new medical and biological advances and more effective allocation policies in the field of organ transplantation.