Estimating residual variance in random forest regression

کد مقاله	کد نشریه	سال انتشار	مقاله انگلیسی	نسخه تمام متن
415889	681253	2011	14 صفحه PDF	دانلود رایگان

عنوان انگلیسی مقاله ISI

دانلود مقاله + سفارش ترجمه

دانلود مقاله ISI انگلیسی

رایگان برای ایرانیان

کلمات کلیدی

Greater male variability hypothesis Proximity measure - اندازه گیری نزدیکی Bootstrap - بوت استرپ Sex differences - تفاوت های جنسیتی Regression tree - درخت رگرسیون Nonparametric regression - رگرسیون ناپارامتری Gender gap - شکاف جنسیتی

موضوعات مرتبط

مهندسی و علوم پایه مهندسی کامپیوتر نظریه محاسباتی و ریاضیات

پیش نمایش صفحه اول مقاله

Estimating residual variance in random forest regression

چکیده انگلیسی

Random forest, a data-mining technique which uses multiple classification or regression trees, is a popular algorithm used for prediction. Inference and goodness-of-fit assessment, however, may require an estimator of variability; in many applications the residual variance is of primary interest. This paper proposes two estimators of residual variance for random forest regression that take advantage of byproducts of the algorithm. The first estimator is based on the residual sum of squares from a random forest fit and uses a bootstrap bias correction. The second estimator is a difference-based estimator that uses proximity measures as weights. The estimators are evaluated through Monte Carlo simulations. Applications of the methods to the problem of assessing the relative variability of males and females on cognitive and achievement tests are discussed, and the methods are applied to estimate the residual variance in test scores for male and female students on the mathematics portion of the 2007 Arizona Instrument to Measure Standards.

ناشر

Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Computational Statistics & Data Analysis - Volume 55, Issue 11, 1 November 2011, Pages 2937–2950

نویسندگان

Guillermo Mendez, Sharon Lohr,

علوم انسانی و هنر

فنی، مهندسی و علوم پایه

پزشکی و سلامت

بیو تکنولوژی

پذیرش سفارش ترجمه

Estimating residual variance in random forest regression

دسترسی سریع

ارتباط

English Website