已入深夜,您辛苦了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!祝你早点完成任务,早点休息,好梦!

From Fixed-X to Random-X Regression: Bias-Variance Decompositions, Covariance Penalties, and Prediction Error Estimation

协变量 数学 普通最小二乘法 协方差 统计 差异(会计) 对比度(视觉) 随机性 协方差分析 线性回归 应用数学 计算机科学 会计 人工智能 业务
作者
Saharon Rosset,Ryan J. Tibshirani
出处
期刊: 卷期号:115 (529): 138-151 被引量:46
标识
DOI:10.1080/01621459.2018.1424632
摘要

In statistical prediction, classical approaches for model selection and model evaluation based on covariance penalties are still widely used. Most of the literature on this topic is based on what we call the "Fixed-X" assumption, where covariate values are assumed to be nonrandom. By contrast, it is often more reasonable to take a "Random-X" view, where the covariate values are independently drawn for both training and prediction. To study the applicability of covariance penalties in this setting, we propose a decomposition of Random-X prediction error in which the randomness in the covariates contributes to both the bias and variance components. This decomposition is general, but we concentrate on the fundamental case of ordinary least-squares (OLS) regression. We prove that in this setting the move from Fixed-X to Random-X prediction results in an increase in both bias and variance. When the covariates are normally distributed and the linear model is unbiased, all terms in this decomposition are explicitly computable, which yields an extension of Mallows' Cp that we call RCp. RCp also holds asymptotically for certain classes of nonnormal covariates. When the noise variance is unknown, plugging in the usual unbiased estimate leads to an approach that we call RCp ^, which is closely related to Sp, and generalized cross-validation (GCV). For excess bias, we propose an estimate based on the "shortcut-formula" for ordinary cross-validation (OCV), resulting in an approach we call RCp+. Theoretical arguments and numerical simulations suggest that RCp+ is typically superior to OCV, though the difference is small. We further examine the Random-X error of other popular estimators. The surprising result we get for ridge regression is that, in the heavily regularized regime, Random-X variance is smaller than Fixed-X variance, which can lead to smaller overall Random-X error. Supplementary materials for this article are available online.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
刚刚
1秒前
迷路博完成签到,获得积分10
1秒前
xin发布了新的文献求助10
3秒前
4秒前
5秒前
充电宝应助bottlemonster采纳,获得10
5秒前
ju8715发布了新的文献求助10
6秒前
竹隐完成签到,获得积分10
7秒前
搜集达人应助zzxiao采纳,获得10
8秒前
yin完成签到,获得积分20
9秒前
9秒前
dawn完成签到,获得积分10
9秒前
重要靳完成签到 ,获得积分10
10秒前
14秒前
15秒前
hedinghong完成签到,获得积分10
16秒前
isasi完成签到,获得积分10
17秒前
QAQ完成签到 ,获得积分10
17秒前
18秒前
18秒前
20秒前
Aveline完成签到 ,获得积分10
21秒前
Akim应助冷静的不言采纳,获得10
21秒前
科研通AI6.4应助ju8715采纳,获得10
21秒前
科研通AI6.4应助syx采纳,获得10
22秒前
汉堡包应助huanglanlan采纳,获得10
22秒前
wqdoctor完成签到,获得积分10
23秒前
清爽的厉发布了新的文献求助30
23秒前
CodeCraft应助质谱仪采纳,获得10
23秒前
慕青应助海绵宝宝采纳,获得10
25秒前
研友_LMyNzL发布了新的文献求助10
25秒前
今后应助SSS采纳,获得10
25秒前
小面面完成签到 ,获得积分10
27秒前
30秒前
Ade完成签到,获得积分10
30秒前
31秒前
CYYDNDB完成签到 ,获得积分10
32秒前
整齐的千万完成签到 ,获得积分10
33秒前
登登完成签到 ,获得积分10
34秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
The anomeric effect 1000
Principles of town planning: translating concepts to applications 1000
1 Peter and Christ's Descent to the Dead in Its Early Christian Reception 700
Organizational Behavior 510
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7732145
求助须知:如何正确求助?哪些是违规求助? 9282903
关于积分的说明 20155312
捐赠科研通 7309456
什么是DOI,文献DOI怎么找? 3303909
关于科研通互助平台的介绍 2456659
邀请新用户注册赠送积分活动 2312950