دانلود رایگان مقاله: درباره استفاده از C Harrell برای پیش بینی خطر بالینی از طریق جنگل های بازمانده تصادفی

کد مقاله	کد نشریه	سال انتشار	مقاله انگلیسی	نسخه تمام متن
383003	660799	2016	10 صفحه PDF	دانلود رایگان

عنوان انگلیسی مقاله ISI

On the use of Harrell’s C for clinical risk prediction via random survival forests

ترجمه فارسی عنوان

درباره استفاده از C Harrell برای پیش بینی خطر بالینی از طریق جنگل های بازمانده تصادفی

دانلود مقاله + سفارش ترجمه

دانلود مقاله ISI انگلیسی

رایگان برای ایرانیان

کلمات کلیدی

شاخص تطابق؛ آنالیز تاریخچه رویداد. آمار لگ رتبه ؛ جنگل های بازمانده تصادفی. پیش بینی خطر؛ قوانین اسپلیت

Event history analysis - تجزیه و تحلیل تاریخ رویداد concordance index - شاخص همبستگی Risk prediction - پیش بینی ریسک

موضوعات مرتبط

مهندسی و علوم پایه مهندسی کامپیوتر هوش مصنوعی

پیش نمایش مقاله

درباره استفاده از C Harrell برای پیش بینی خطر بالینی از طریق جنگل های بازمانده تصادفی

چکیده انگلیسی

• Harrell’s C is proposed as a split criterion in random survival forests.
• Split points of continuous predictor variables differ substantially between Harrell’s C and log-rank splitting.
• The log-rank statistic has a stronger end-cut preference than Harrell’s C.
• Harrell’s C outperforms log-rank splitting in smaller scale studies.
• Harrell’s C outperforms log-rank splitting if the censoring rate is high.

Random survival forests (RSF) are a powerful method for risk prediction of right-censored outcomes in biomedical research. RSF use the log-rank split criterion to form an ensemble of survival trees. The most common approach to evaluate the prediction accuracy of a RSF model is Harrell’s concordance index for survival data (‘C index’). Conceptually, this strategy implies that the split criterion in RSF is different from the evaluation criterion of interest. This discrepancy can be overcome by using Harrell’s C for both node splitting and evaluation. We compare the difference between the two split criteria analytically and in simulation studies with respect to the preference of more unbalanced splits, termed end-cut preference (ECP). Specifically, we show that the log-rank statistic has a stronger ECP compared to the C index. In simulation studies and with the help of two medical data sets we demonstrate that the accuracy of RSF predictions, as measured by Harrell’s C, can be improved if the log-rank statistic is replaced by the C index for node splitting. This is especially true in situations where the censoring rate or the fraction of informative continuous predictor variables is high. Conversely, log-rank splitting is preferable in noisy scenarios. Both C-based and log-rank splitting are implemented in the R package ranger. We recommend Harrell’s C as split criterion for use in smaller scale clinical studies and the log-rank split criterion for use in large-scale ‘omics’ studies.

ناشر

Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Expert Systems with Applications - Volume 63, 30 November 2016, Pages 450–459

نویسندگان

Matthias Schmid, Marvin N. Wright, Andreas Ziegler,

علوم انسانی و هنر

فنی، مهندسی و علوم پایه

پزشکی و سلامت

بیو تکنولوژی

پذیرش سفارش ترجمه

دانلود رایگان مقاله ISI : درباره استفاده از C Harrell برای پیش بینی خطر بالینی از طریق جنگل های بازمانده تصادفی

دسترسی سریع

ارتباط

English Website