Double Descent and Emergent Smoothing in Model Averaging Prediction

This paper investigates the predictive performance of model averaging in high-dimensional linear regression where the number of regressors is comparable to the sample size. We demonstrate that the double descent trajectory manifests within the model averaging framework, where the ensemble inherits the variance explosion of individual models near the interpolation boundary. However, we reveal that weighted aggregation simultaneously triggers an emergent smoothing effect that structurally suppresses the localized risk divergence, indicating that strategic weight choice serves as a vital stabilizing mechanism. Leveraging tools from random matrix theory, we derive the exact limiting out-of-sample risk under a nested model setting and provide a comprehensive characterization of the risk landscape. Building on these asymptotic results, we propose the Large Model Averaging (LaMA) method, which introduces a novel criterion incorporating in-sample bias and asymptotic out-of-sample variance to balance fitting accuracy and generalization. Numerical studies and real data applications confirm that LaMA achieves superior predictive accuracy in high-dimensional environments.

翻译：本文研究回归变量数量与样本量相当的高维线性回归中模型平均的预测性能。我们证明双重下降轨迹在模型平均框架中显现，当集成模型继承个体模型在插值边界附近的方差爆炸特性时。然而，我们揭示加权聚合同时会触发一种涌现平滑效应，该效应从结构上抑制了局部风险发散，表明策略性权重选择可作为关键稳定机制。利用随机矩阵理论工具，我们在嵌套模型设定下推导出精确的渐近外样本风险，并提供风险景观的全面表征。基于这些渐近结果，我们提出大模型平均方法（LaMA），该方法引入融合样本内偏差与渐近外样本方差的新型准则，以平衡拟合精度与泛化能力。数值实验与真实数据应用证实LaMA在高维环境中能实现更优的预测精度。

相关内容

MoDELS

关注 45

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

【AAAI2026】《SimDiff：用于时间序列点预测的更简单但更优的扩散模型》

专知会员服务

14+阅读 · 2025年11月25日

用于多模态对齐的基础模型表征潜力：一项综述

专知会员服务

18+阅读 · 2025年10月8日

【博士论文】Stein变分梯度下降与基于共识的优化：趋向于收敛分析与泛化，195页pdf

专知会员服务

20+阅读 · 2024年6月2日

大型语言模型在预测和异常检测中的应用综述

专知会员服务

70+阅读 · 2024年2月19日