---
格式版本: 2
标题: "Gradient-Update Mismatch: Rethinking Conflict-Free Training of Physics-Informed Neural Networks"
原文链接: "https://arxiv.org/abs/2609.01558"
发布日期: "2026-09-01"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Tue, 1 Sep 2026 17:20:13 UTC (696 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-09-02T18:52:20+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-09-02T18:51:39+08:00"
入库时间: "2026-09-02T10:52:20.729Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=Scale-up&searchtype=all"
匹配关键词:
  - "Scale-up"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 26
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文主线是物理信息神经网络的梯度冲突与优化器更新对齐，提出GUM与GUA并报告误差和冲突率实验，不涉及超节点、AI Rack、机架级互连、供电、散热或RAS。来源为可追溯的arXiv预印本摘要页；固定知识库未见该方法的历史重复，但其新增算法和实验仅属于模型训练优化，与超节点业务无关，也无产品、客户、量产或部署信号。命中应用与模型效率/非机架级基础设施研究否决项。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-09-02T18:53:20+08:00"
AI主题相关性: 0
AI来源权威性: 11
AI新颖性: 7
AI技术细节: 0
AI商业部署信号: 0
AI完整性: 8
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.81"
AI评分知识库SHA256: "3b93d12e47749b3512f545f51c44c011bdc0931677c2cfe61e4df59a2b1a5a48"
AI评分知识库检索词: "[\"Scale-up\",\"PDF\",\"arxiv.org/pdf/2609.01558\",\"HTML\",\"arxiv.org/html/2609.01558v1\",\"PINNs\",\"GUM\",\"GUA\",\"PINN\",\"L_2\",\"URL\",\"github.com/JingXiao10/GUA\"]"
AI评分知识库命中: "[{\"id\":\"runtime-569639751a0dbe7ef3ffccc5\",\"title\":\"[2608.17503] Predict Before Replay: Joint FEC and Flight Control for Reliable Scale-Up Links\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-18\",\"matchedTerms\":[\"Scale-up\",\"PDF\",\"HTML\",\"URL\"],\"rank\":-12.235265611289746},{\"id\":\"july-correct-0111\",\"title\":\"StrataCL: Fabric-Native Communication Library for Production Supernodes\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"PDF\",\"HTML\",\"URL\"],\"rank\":-8.619329601607559},{\"id\":\"runtime-ffe0e6dc6f3bdd3f1ce31ae8\",\"title\":\"5000亿美元！英伟达押注AI基础设施 | SDNLAB | 专注网络创新技术\",\"sourceType\":\"ai_excellent_article\",\"time\":\"\",\"matchedTerms\":[\"Scale-up\",\"HTML\",\"URL\"],\"rank\":-6.887458691269215},{\"id\":\"july-correct-0001\",\"title\":\"全球首颗2nm GPU来了！苏姿丰甩出“最强AI机架”，CPU性能干翻英伟达 - 智东西\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"HTML\"],\"rank\":-6.553580658270128},{\"id\":\"july-correct-0081\",\"title\":\"AAI 2026: 6th Gen AMD EPYC Server CPUs Power the Agentic Data Center\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"PDF\",\"HTML\",\"URL\"],\"rank\":-6.004365679566843}]"
AI摘要: "该研究提出物理信息神经网络训练中的“梯度更新失配”（GUM）问题，指优化器可能破坏梯度手术构造的无冲突方向；并据此提出“梯度更新对齐”（GUA）方法，将优化器更新投影回无冲突锥内。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-02T17:13:26.620Z"
采集批次: "2026年9月2日18点28分48秒"
采集批次ID: "20260902-182848-894"
去重键: "https://arxiv.org/abs/2609.01558"
---

## Computer Science > Machine Learning

## Title:Gradient-Update Mismatch: Rethinking Conflict-Free Training of Physics-Informed Neural Networks

Authors:[Jing Xiao](https://arxiv.org/search/cs?searchtype=author&query=Xiao,+J), [Xinhai Chen](https://arxiv.org/search/cs?searchtype=author&query=Chen,+X), [Qinglin Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+Q), [Menghan Jia](https://arxiv.org/search/cs?searchtype=author&query=Jia,+M), [Zhiquan Lai](https://arxiv.org/search/cs?searchtype=author&query=Lai,+Z), [Dongsheng Li](https://arxiv.org/search/cs?searchtype=author&query=Li,+D), [Jie Liu](https://arxiv.org/search/cs?searchtype=author&query=Liu,+J), [Tiejun Li](https://arxiv.org/search/cs?searchtype=author&query=Li,+T)

[View PDF](https://arxiv.org/pdf/2609.01558) [HTML (experimental)](https://arxiv.org/html/2609.01558v1)

> Abstract:Training Physics-Informed Neural Networks (PINNs) requires jointly optimizing physics residual and initial/boundary condition loss terms, which often induce conflicting gradients. Gradient surgery methods mitigate this issue by constructing directions from loss-specific gradients to reduce conflict before optimizer transformation. However, even when the constructed direction is conflict-free, this property may not be preserved after optimizer transformation. Let $a_t$ denote the direction constructed by gradient surgery, $u_t$ the optimizer proposal, and $\mathcal{C}_t$ the conflict-free cone induced by the loss-specific gradients. We show that modern optimizers can transform $a_t$ through mechanisms such as historical state, adaptive scaling, preconditioning, or decoupled weight decay, so $a_t \in \mathcal{C}_t$ does not generally imply $u_t \in \mathcal{C}_t$. We refer to this optimizer-induced discrepancy in conflict-freeness between $a_t$ and $u_t$ as Gradient-Update Mismatch (GUM). Accordingly, we propose Gradient-Update Alignment (GUA), which projects $u_t$ onto $\mathcal{C}_t$ to obtain the aligned update $p_t$ and applies $p_t$ to the parameters. When the optimizer maintains internal state, GUA further adjusts this state toward targets reconstructed from the applied update. We conduct extensive experiments and find that GUM is widespread across momentum, adaptive, and curvature-based optimizers, with conflict rates reaching up to 86.3%. Across all PINN settings, GUA achieves conflict-free applied updates and consistently improves various gradient surgery methods, reducing the relative $L_2$ error by up to 98.2% in individual settings. Data and code are available at [this https URL](https://github.com/JingXiao10/GUA).

| Subjects: | Machine Learning (cs.LG) |
| --- | --- |
| Cite as: | [arXiv:2609.01558](https://arxiv.org/abs/2609.01558) \[cs.LG\] |
|  | (or [arXiv:2609.01558v1](https://arxiv.org/abs/2609.01558v1) \[cs.LG\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2609.01558](https://doi.org/10.48550/arXiv.2609.01558) |

## Submission history

From: Xinhai Chen \[[view email](https://arxiv.org/show-email/46c678a5/2609.01558)\]  
**\[v1\]** Tue, 1 Sep 2026 17:20:13 UTC (696 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2609.01558) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
