---
格式版本: 2
标题: "Self-Bounding Regret Matching+ in Potential Games and Product-Simplex Optimization"
原文链接: "https://arxiv.org/abs/2608.17417"
发布日期: "2026-08-18"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Tue, 18 Aug 2026 06:36:53 UTC (26 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-20T14:43:02+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-20T14:38:42+08:00"
入库时间: "2026-08-20T06:43:02.415Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=Oracle&searchtype=all"
匹配关键词:
  []
相关厂家:
  - "Oracle"
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 5
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文为博弈论与优化算法论文，与超节点、AI Rack、机柜级AI基础设施、供电散热互连等主题完全无关，无任何相关技术或商业信息。"
AI质检模型: "ali-deepseek-v4-flash"
AI质检时间: "2026-08-20T14:46:12+08:00"
AI主题相关性: 0
AI来源权威性: 5
AI新颖性: 0
AI技术细节: 0
AI商业部署信号: 0
AI完整性: 5
AI摘要: "该文为后悔匹配+（RM+）算法导出一条精确的单步守恒律，证明其遗憾由中心化时间变异控制，并据此在有限精确势博弈的交替、懒惰及循环玩法下达成均匀有界遗憾和ε平方复杂度；"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-07T03:21:50.005Z"
采集批次: "2026年8月20日14点19分32秒"
采集批次ID: "20260820-141932-079"
去重键: "https://arxiv.org/abs/2608.17417"
---

## Computer Science > Computer Science and Game Theory

## Title:Self-Bounding Regret Matching+ in Potential Games and Product-Simplex Optimization

[View PDF](https://arxiv.org/pdf/2608.17417) [HTML (experimental)](https://arxiv.org/html/2608.17417v1)

> Abstract:Regret matching+ (RM+) is parameter free, scale invariant, and central to large game solving, but its only general individual-regret guarantee grows as $\sqrt{T}$. A recent ICLR result used this envelope to prove that RM+ reaches an $\epsilon$ -stationary point of a smooth objective over a product of simplices in $O(\epsilon^{-4})$ iterations, or $O(\epsilon^{-8})$ from the standard zero initialization. We give an exact one-step conservation law for RM+. It states that forward utility gain pays for both squared state motion and growth of the regret-state norm. Norm growth is at most $\sqrt{m-1}$ times forward gain for $m$ actions, and the coefficient is sharp. This yields four results for unmodified RM+. Its regret on any utility path is controlled by centered temporal variation. Its regret is uniformly bounded under alternating play in every finite exact potential game, resolving an open question and making squared activation gaps summable. Both certified lazy and ordinary cyclic play attain an $\epsilon^{-2}$ exponent. On any smooth, possibly nonconcave simplex objective, RM+ finds an $\epsilon$ -KKT point in $O(\epsilon^{-2})$ iterations. Most broadly, for a smooth objective over an arbitrary product of simplices, cyclic block RM+ attains the same $O(\epsilon^{-2})$ exponent from arbitrary initialization, with an explicit trajectory-dependent constant. The proof controls the finite objective loss caused by low-state blocks and then self-bounds every block state and the total squared path length. Complete proofs cover zero states, sharpness, common-profile stationarity, and robust gain dominance. Oracle-normalized diagnostics compare RM+ with predictive and smooth extra-gradient variants on graphical potential games and dense nonconvex objectives.

| Subjects: | Computer Science and Game Theory (cs.GT) |
| --- | --- |
| Cite as: | [arXiv:2608.17417](https://arxiv.org/abs/2608.17417) \[cs.GT\] |
|  | (or [arXiv:2608.17417v1](https://arxiv.org/abs/2608.17417v1) \[cs.GT\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.17417](https://doi.org/10.48550/arXiv.2608.17417) |

## Submission history

From: Subhashini Jayawardhana \[[view email](https://arxiv.org/show-email/719a2f6d/2608.17417)\]  
**\[v1\]** Tue, 18 Aug 2026 06:36:53 UTC (26 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.17417) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
