نوع مقاله : مقاله پژوهشی
نویسندگان
1
دانشجوی دکتری تخصصی، انفورماتیک پزشکی، کمیته تحقیقات دانشجویی، دانشگاه علوم پزشکی مشهد، مشهد، ایران
2
دانشجوی دکتری تخصصی، انفورماتیک پزشکی، گروه پژوهشی انفورماتیک سرطان، مرکز تحقیقات سرطان پستان جهاد دانشگاهی، تهران، ایران
3
دانشجوی دکتری تخصصی، مدیریت اطلاعات سلامت، مرکز تحقیقات علوم مدیریت و اقتصاد سلامت، دانشکدهی مدیریت و اطلاعرسانی پزشکی، دانشگاه علوم پزشکی ایران، تهران، ایران
4
استادیار، انفورماتیک پزشکی، دانشکدهی پزشکی، دانشگاه علوم پزشکی مشهد، مشهد، ایران
چکیده
مقدمه: در دههی اخیر الگوریتمهای یادگیری ماشین به ابزار مفیدی جهت دادهکاوی در دادههای پزشکی، برای تولید مدلهای پیشبینی تبدیل شدهاند. سوختگی از جمله بیماریهایی است که پیشبینی پیامد آن از اهمیت زیادی برخوردار است. هدف این مطالعه بررسی عملکرد دو الگوریتم پراستفادهی یادگیری ماشین یعنی شبکهی عصبی و درخت تصمیم و مقایسه با روش آماری رگرسیون لجستیک در پیشبینی پیامد بیماران سوختگی بوده است. روش بررسی: در این مطالعه مشاهدهای گذشتهنگر، پس از انجام پردازش اولیهی دادهها و تعیین پیامد (زنده یا فوت)، دو الگوریتم یادگیری ماشین (شبکهی عصبی و درخت تصمیم) به همراه روش آماری رگرسیون لجستیک برای تولید مدلهای پیشبینی روی دادههای 4804 بیمار سوختگی بیمارستان طالقانی اهواز مربوط به سالهای 1380 تا 1386 اعمال گردید. برای پردازش اولیهی دادهها نرمافزار SPSS16 و در مرحلهی مدلسازی از Clementine 12.0 استفاده شد. همچنین با بهکارگیری تکنیک 10-Fold Cross Validation، معیارهای ارزیابی کارایی برای دادههای تست محاسبه و مقایسه شدند. یافتهها: نتایج نشان داد الگوریتم شبکهی عصبی با دقت 97 درصد منجر به دقیقترین مدل روی دادههای مورد مطالعه میشود. مدل درخت تصمیم با دقت 95 درصد در ردهی دوم و مدل رگرسیون لجستیک با دقت 90 درصد کمترین دقت را داشت. سایر معیارهای ارزیابی مانند حساسیت (Sensitivity)، ویژگی (Specificity)، PPV (Positive Predictive Value) و NPV (Negative Predictive Value) و AUC (Area Under the Curve) نیز کارایی مدل شبکهی عصبی را بالاتر از دو مدل دیگر نشان دادند. نتیجهگیری: تحلیل نتایج این مطالعه و مطالعات مشابه نشان میدهند که الگوریتمهای یادگیری ماشین نسبت به روشهای آماری منجر به تولید مدلهای دقیقتری میشوند. بسته به ماهیت و میزان دادهها و همچنین جامعهی پژوهش، الگوریتمهای مختلف یادگیری ماشین، رفتارهای متفاوتی دارند که بهنظر میرسد دقت مدلهای شبکهی عصبی از سایر مدلها بیشتر میباشد. واژههای کلیدی: دادهکاوی؛ یادگیری ماشین؛ پیشبینی؛ درخت تصمیم؛ شبکهی عصبی مصنوعی؛ سوختگیها
کلیدواژهها
عنوان مقاله English
Using Data Mining to Predict Outcome in Burn Patients: A Comparison Between Several Algorithms
نویسندگان English
Ehsan Nabovati
1
Amir Abas Azizi
1
Ebrahim Abbasi
2
Hassan Vakili-Arki
1
Javad Zarei
3
Amir Reza Razavi
4
1
PhD Candidate, Medical Informatics, Student Research Committee, Mashhad University of Medical Sciences, Mashhad, Iran
2
PhD Candidate, Medical Informatics, Cancer Informatics Research Group, BCRC(Breast Cancer Research Center), ACECR (Academic Center for Education, Culture and Research), Tehran, Iran
3
PhD Candidate, Health Information Management, Health Management and Economics Research Center, School of Health Management and Information Sciences, Iran University of Medical Sciences, Tehran, Iran
4
Assistant Professor, Medical Informatics, Mashhad University of Medical Sciences, Mashhad, Iran
چکیده English
Introduction: In the past decades, machine learning algorithms have become a useful tool for data mining within huge amounts of health data to create prediction models. Burn is one of the diseases that predicting of its outcome has high importance. The aim of this study was to survey two widely used machine learning algorithms; neural network and decision tree, and compare them with logistic regression method to predict the outcome of burn patients. Methods: In this retrospective observational study, following preprocessing of the data and determining the outcome of patient (live or death), two well-known machine learning algorithms (neural network and decision tree) and logistic regression method were used to create prediction models using data from 4804 burn patients hospitalized in Taleghani Burn Center in Ahvaz during the years 2001-2007. The preprocessing of the data was performed using SPSS (Version16.0), and in the modeling phase, Clementine (Version 12.0) software was used. Moreover, 10-fold cross validation technique was used to validate the model and criteria for evaluating the performance of models were measured and compared. Results: The results showed that the neural network algorithm with accuracy of 97% resulted the most accurate model on the studied data. The decision tree model with 95% accuracy was in the second place and the logistic regression model with an accuracy of 90% was the least accurate. Moreover other evaluating criteria such as sensitivity, specificity, PPV, NPV and AUC showed that performance of the neural network model was better than the others. Conclusion: The current study shows that machine learning algorithms compared with statistical methods create more accurate models. In analyzing the current data, the model created by artificial neural network is more accurate than the other machine learning algorithm, decision tree. Keywords: Data Mining; Machine Learning; Forecasting; Decision Tree; Artificial Neural Network; Burns
کلیدواژهها English
Data Mining
Machine Learning
Forecasting
Decision Tree
Artificial Neural Network
Burns