---
格式版本: 2
标题: "When Does Dynamic Ensembling Pay Off? Diagnosing Regionwise Gains in Regression under Distribution Shift"
原文链接: "https://arxiv.org/abs/2608.18330"
发布日期: "2026-08-18"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Tue, 18 Aug 2026 21:31:30 UTC (902 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-20T15:52:41+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-20T15:48:16+08:00"
入库时间: "2026-08-20T07:52:41.290Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=model%20fit&searchtype=all"
匹配关键词:
  - "model fit"
  - "deployment"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 5
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文为arXiv机器学习动态集成回归论文，与超节点/AI Rack/机柜级AI基础设施完全无关，仅命中model fit搜索词，无任何相关技术或商业信息。"
AI质检模型: "ali-deepseek-v4-flash"
AI质检时间: "2026-08-20T15:59:15+08:00"
AI主题相关性: 0
AI来源权威性: 5
AI新颖性: 0
AI技术细节: 0
AI商业部署信号: 0
AI完整性: 0
AI摘要: "研究者提出 D_CF5 估计器，用少量目标域探针判断区域级动态集成相比最佳"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-07T03:21:24.192Z"
采集批次: "2026年8月20日14点19分32秒"
采集批次ID: "20260820-141932-079"
去重键: "https://arxiv.org/abs/2608.18330"
---

## Computer Science > Machine Learning

## Title:When Does Dynamic Ensembling Pay Off? Diagnosing Regionwise Gains in Regression under Distribution Shift

Authors:[Tianxin Zhou](https://arxiv.org/search/cs?searchtype=author&query=Zhou,+T), [Ruixi Lin](https://arxiv.org/search/cs?searchtype=author&query=Lin,+R)

[View PDF](https://arxiv.org/pdf/2608.18330) [HTML (experimental)](https://arxiv.org/html/2608.18330v1)

> Abstract:Whether input-dependent ("dynamic") combination of a regression model pool beats the best static blend depends on the shift and is rarely known before deployment. Can a small labeled target-domain probe tell us when reallocating trust across regions of the input space will pay off? We answer this with $\widehat{D}_{\mathrm{CF5}}$, which estimates from the probe the cross-fitted gain of the regionwise convex combination over the best static convex blend: the realizable value of deciding, region by region, whom to trust. Across a frozen suite of 12 dataset-shift pairs (spatial, temporal, domain, feature-cluster), $\widehat{D}_{\mathrm{CF5}}$ predicts realized regionwise test gains with dataset-level Spearman $+0.98$ (95% CI $\[+0.83, +1.00\]$; $p=5\times10^{-5}$), including two cases overturning preregistered expectations. The relationship holds in a 16-pair sensitivity analysis (Spearman $+0.83$), whereas alternative probe diagnostics reach at most $+0.66$. This contrast isolates regional trust reallocation: correlation is $+0.98$ for regionwise-convex gain, but $+0.01$ for smooth covariate-dependent stacking after affine correction. A controlled generator shows dynamic gains arise from the interaction of shift heterogeneity and local competence, increase with shift severity, and become realizable between 128 and 256 probe labels in the tested grid. The Probe-Validated Ensemble Selector chooses among a static affine stacker and dynamic realizers, deploying a candidate only when a held-out lower confidence bound clears the static-convex floor. In a preregistered prospective batch, it matched or improved the floor in all 12 runs; two deployments reduced test risk by 11% and 16%, while the gate rejected a candidate whose un-gated deployment incurred $>30\times$ the static loss. We release OpenRegShift, a reproducible evaluation harness for regression ensembles under distribution shift.

| Comments: |  |
| --- | --- |
| Subjects: | Machine Learning (cs.LG) |
| Cite as: | [arXiv:2608.18330](https://arxiv.org/abs/2608.18330) \[cs.LG\] |
|  | (or [arXiv:2608.18330v1](https://arxiv.org/abs/2608.18330v1) \[cs.LG\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.18330](https://doi.org/10.48550/arXiv.2608.18330) |

## Submission history

From: Ruixi Lin \[[view email](https://arxiv.org/show-email/792e8551/2608.18330)\]  
**\[v1\]** Tue, 18 Aug 2026 21:31:30 UTC (902 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.18330) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
