---
格式版本: 2
标题: "What NVHBM Means: The Contest for Power Over HBM Begins"
原文链接: "https://www.damnang.com/p/what-nvhbm-means-the-contest-for"
发布日期: "2026-08-30"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "llm:scrape:extracted_text"
发布时间证据: "第 8 行：Aug 30, 2026"
发布时间校准原因: "正文中靠近标题的位置明确显示 Aug 30, 2026，且与 provider publishedAt 和 datePublished 一致，属于文章发布时间标注。"
发布时间校准置信度: "1"
发布时间候选数量: 15
发布时间严格候选数量: 2
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-31T11:56:23+08:00"
发布时间仲裁状态: "confirmed"
发布时间仲裁尝试次数: 3
发布时间仲裁耗时毫秒: 23592
发现时间: "2026-08-31T11:48:50+08:00"
入库时间: "2026-08-31T03:57:11.944Z"
来源平台: "Substack 数据中心相关博客搜索"
搜索渠道: "source_template"
搜索词: "site:substack.com HBM"
匹配关键词:
  - "HBM"
  - "GPU"
  - "XPU"
  - "Scale-up"
  - "NVLink"
  - "performance"
  - "bandwidth"
  - "AI"
相关厂家:
  - "NVIDIA"
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
图片摘要:
  - "✗ ./assets/img-b53e97f3.webp | decorative | 装饰性图片，与正文主题无关"
  - "★ ./assets/img-8e41d035.webp | diagram | 对比sHBM与cHBM的截面结构差异，展示NVHBM在base die集成NVIDIA控制器与接口的位置关系"
  - "★ ./assets/img-857bf4bd.webp | infographic | 对比sHBM、cHBM与NVHBM三者的价值构成分布，直观体现设计权从存储厂商向NVIDIA平台层转移"
  - "★ ./assets/img-477d9be9.webp | diagram | 图示NVIDIA在NVHBM架构中定义的控制器与接口范围，明确其与存储厂商供应部分的分界"
AI优质: "否"
AI打分: 68
AI分档: "召回候选"
AI质检状态: "不通过"
AI打分理由: "正文主线是NVHBM如何把内存控制器移入HBM基底裸片，并影响XPU面积、功耗、供应商切换成本及NVIDIA对机架级AI平台的控制，技术分析较深入。当前来源为独立Substack博客，虽链接官方原文，但并非一手发布。历史知识库同日官方稿已完整披露NVHBM、多供应商、带宽/功耗/面积收益、NVLink Fusion及Amazon Trainium4合作等核心事实；本文主要新增bespoke与platform-defined cHBM的价值分配分析及12-Hi转8-Hi的1.5倍理论推演，缺少新的实测、客户部署、量产节点或独立供应链证据。页面本身具有分析参考价值，但未明确提供足以进入高质量库的新增可核验事件，且发布日期缺失，因此作为召回候选，不通过最终准入。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-08-31T11:57:27+08:00"
AI主题相关性: 18
AI来源权威性: 9
AI新颖性: 9
AI技术细节: 17
AI商业部署信号: 7
AI完整性: 8
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.62"
AI评分知识库SHA256: "b53abf2242d71cbdd0daf92240736320626be4a6f0e164c2fafe18c1a554a4a3"
AI评分知识库检索词: "[\"NVIDIA\",\"HBM\",\"site:substack.com HBM\",\"NVLink\",\"HBM3E\",\"HBM4\",\"rack-scale\",\"RAS\",\"GPU\",\"XPU\",\"NVHBM\",\"HBM4E\"]"
AI评分知识库命中: "[{\"id\":\"runtime-676f90ee5f88ebab3b504077\",\"title\":\"NVIDIA NVLink Fusion Brings NVHBM to Next-Generation AI Infrastructure\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-26\",\"matchedTerms\":[\"NVIDIA\",\"HBM\",\"NVLink\",\"HBM4\",\"rack-scale\",\"RAS\",\"GPU\",\"XPU\",\"NVHBM\",\"HBM4E\"],\"rank\":-34.06981283704393},{\"id\":\"runtime-52343a012216677e078ffdd2\",\"title\":\"NVIDIA NVLink Fusion Expands With NVHBM Custom High-Bandwidth Memory\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-26\",\"matchedTerms\":[\"NVIDIA\",\"HBM\",\"NVLink\",\"HBM4\",\"rack-scale\",\"RAS\",\"GPU\",\"XPU\",\"NVHBM\",\"HBM4E\"],\"rank\":-31.957205169997252},{\"id\":\"july-correct-0095\",\"title\":\"Inside NVIDIA Rubin GPU Architecture: Powering the Era of Agentic AI\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"NVIDIA\",\"HBM\",\"NVLink\",\"HBM3E\",\"HBM4\",\"rack-scale\",\"RAS\",\"GPU\"],\"rank\":-16.672648890183897},{\"id\":\"runtime-4abbccc42d96af674efc7768\",\"title\":\"OpenAI’ Jalapeño: Better Than Nvidia Blackwell\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-25\",\"matchedTerms\":[\"NVIDIA\",\"HBM\",\"HBM3E\",\"HBM4\",\"rack-scale\",\"RAS\",\"GPU\",\"XPU\"],\"rank\":-11.882053391589572},{\"id\":\"july-correct-0020\",\"title\":\"NVIDIA Vera Rubin：引領代理 AI 的時代\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"NVIDIA\",\"HBM\",\"NVLink\",\"RAS\",\"GPU\"],\"rank\":-11.663836413688053}]"
AI摘要: "NVIDIA于2026年8月26日发布NVHBM规范，将自定义内存控制器放进HBM基底die，并让多家内存供应商按同一规格供货。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-08-31T23:37:19.622Z"
采集批次: "2026年8月31日11点48分36秒"
采集批次ID: "20260831-114836-897"
去重键: "https://www.damnang.com/p/what-nvhbm-means-the-contest-for"
---

## 1\. NVHBM Is Not a New HBM. It Is a Shift in Who Owns the Design

NVIDIA disclosed NVHBM on August 26, 2026. It said it would put a custom memory controller into the HBM base die and let multiple memory providers supply the same NVHBM specification. Versus standard HBM4E it cited up to 30% higher stack bandwidth, up to 15% lower HBM power and up to 25% more XPU compute die area.

![Decorative image.](./assets/img-b53e97f3.webp)

source: https://developer.nvidia.com/blog/nvidia-nvlink-fusion-brings-nvhbm-to-next-generation-ai-infrastructure/

People outside the semiconductor industry sometimes read the name and assume NVHBM is a new form of HBM.

**It is not.**

The industry calls the HBM we normally refer to standard HBM, or sHBM, and calls the products whose base die is tailored to a customer’s requirements cHBM, or custom HBM. Research and development on cHBM has been under way for years. NVHBM is simply the cHBM NVIDIA uses.

The reason NVHBM deserves attention is not that it is technically special.

It is that NVIDIA defines its own cHBM specification, namely the controller and interface that sit in the base die, and intends to have multiple memory suppliers build to that same structure. In other words, the domain memory vendors used to design together with their customers moves wholesale into NVIDIA’s platform layer. That is where its significance lies. Section 4 onward covers this in detail. First, though, we look at why cHBM emerged and what it means for memory vendors.

---

## 2\. The Faster HBM Gets, the More the XPU Pays

HBM has raised bandwidth in two ways: widening the interface between XPU and HBM, and running each signal faster. The interface went from 1,024-bit in HBM3E to 2,048-bit in HBM4, and pin speed keeps climbing.

The problem is that HBM does not bear the cost of higher bandwidth alone. A wider interface requires more PHY and I/O circuitry on the XPU as well, and consumes more die edge and more interposer routing. Higher signal rates add power and signal-integrity burden on top.

In an AI accelerator this cost matters more. The XPU’s leading-edge silicon is an expensive resource that should go to compute and cache wherever possible. Yet the more HBM bandwidth rises, the larger the share of that area and power budget the memory interface takes.

So the problem after HBM4 is not solved by making DRAM itself faster.

What is needed is

***an architecture that delivers the same bandwidth using less XPU area and power**.*

That is where the value of an advanced-logic base die and a custom interface comes from.

---

## 3\. cHBM Reduces That Cost While Creating a New Source of Revenue and Profit

The main disclosed direction for cHBM is to move part of the memory-related logic out of the XPU and into an advanced-logic base die, and to redefine the XPU-to-HBM interface to fit the system. That frees some of the area the HBM PHY and controller occupied on the XPU, and lets the recovered silicon and power budget go back to compute or cache.

![Cross-section comparison of sHBM and cHBM](./assets/img-8e41d035.webp)

Conceptual diagram. Layout and proportions differ from actual devices. The base die does not grow the way it appears here in cHBM. Its size is set by the DRAM core dies stacked above it, and it is drawn enlarged here to show the blocks inside.

For memory vendors this creates a new value pool. In standard HBM, DRAM die, stacking, test and qualification accounted for most of the product value. In cHBM, base-die design, controller and PHY IP, verification, NRE and customer-specific optimization can be added on top.

Micron’s stated expectation that a customized HBM4E base logic die will carry a higher gross margin than standard HBM4E points the same way. cHBM is not simply faster HBM. It is a product through which memory vendors can sell more engineering value.

---

## 4\. NVHBM Can Redistribute cHBM’s Value Capture Toward NVIDIA

> Throughout this article, custom HBM whose architecture is designed anew for each customer is called **bespoke cHBM**, and the model in which a platform owner such as NVIDIA sets the spec and multiple suppliers build the same structure is called **platform-defined cHBM**.

In bespoke cHBM, base-die architecture and qualification can differ by supplier. Once controller, PHY and base die have been designed with one memory vendor, switching to another requires redesign and revalidation. That switching cost is an important reason memory vendors can defend custom design premium and customer relationships.

NVHBM changes this structure. Instead of memory vendors building different cHBM architectures with each customer, NVIDIA defines the controller and interface and multiple memory providers supply the same NVIDIA implementation.

That also changes where the new logic value in an HBM stack accrues. In bespoke cHBM, a memory vendor can go deeper into base-die design, controller, PHY and customer-specific logic and capture proprietary design value. Under NVHBM, NVIDIA defines much of the architecture, so memory vendor differentiation is likely to shift toward DRAM performance, stacking, yield, power and manufacturing execution.

![Value composition across sHBM, cHBM and NVHBM](./assets/img-857bf4bd.webp)

Value composition across sHBM, cHBM and NVHBM

Multi-sourcing reinforces the shift. When several suppliers meet the same NVHBM specification, NVIDIA can qualify suppliers against more common criteria and allocate volume accordingly.

That does not mean memory vendors become fully interchangeable or that prices automatically fall. HBM still differs meaningfully by supplier in capacity, yield, speed, power and thermal behavior, and supply itself remains constrained.

Even so, a common architecture works **in the direction of lowering supplier-specific switching barriers**. The net effect is that NVHBM increases NVIDIA’s influence over custom memory architecture and supplier selection, and can move part of the design value memory vendors could capture in bespoke cHBM toward the platform owner.

---

## 5\. What NVIDIA Is Optimizing Is Not HBM but the Whole AI System

Taken separately, density despec and NVHBM look like HBM optimization. The scope NVIDIA is drawing is wider than that: binding compute, memory and interconnect into a single AI factory architecture rather than treating them as separate components.

Density despec lowers memory capacity per accelerator but lets a constrained DRAM supply produce more HBM stacks. All else equal, moving from 12-Hi to 8-Hi raises the number of stacks buildable from the same DRAM die output by a theoretical factor of 1.5. At the same time, demand for base-die units, one per stack, rises.

In NVHBM, NVIDIA directly defines the architecture between XPU and HBM. Outside that, NVLink and NVSwitch handle Scale-Up while Spectrum-X and InfiniBand handle Scale-Out. NVIDIA has added Scale-Across, Context Memory and BlueField-4-based Scale-In to frame the AI factory as five infrastructure pillars. As rack and data center scale grow, optical connectivity grows in importance alongside them.

Add NVLink Fusion and NVIDIA does not need to build every piece of compute silicon itself. Custom XPUs and CPUs are allowed, while memory, the scale-up fabric and the MGX rack architecture still connect into the NVIDIA platform. Amazon Annapurna Labs being named the first NVHBM collaborator, with NVLink Fusion support signaled from Trainium4, shows this strategy is aimed at custom XPUs beyond NVIDIA’s own GPUs.

![The scope of architecture NVIDIA defines](./assets/img-477d9be9.webp)

The scope of architecture NVIDIA defines

What NVIDIA is expanding is not just GPU share but **the scope of the AI system whose architecture it defines**. From the GPU out to HBM, rack-scale interconnect, networking, context memory and infrastructure access, that boundary keeps widening. NVHBM is one step in bringing the memory layer inside the NVIDIA platform.

---

That is how the market reads NVHBM today: NVIDIA takes the spec and keeps widening the scope it defines. Most commentary stops here.

So are memory vendors simply losing ground? No.

**This is not the last round.** Memory vendors are already preparing for the next one, and in that round there is one position left where what NVIDIA has taken can be won back. It is a domain that cannot be pinned down in a specification, which is exactly why it resists commoditization. Look at where Samsung and SK hynix are putting people and money right now and the direction becomes visible.

The same shift also does not land on all three memory vendors the same way. Set a few conditions, run the numbers per accelerator, and a variable emerges that moves profit far more than design ownership itself. How sensitive that variable is, and how far down the value chain the change carries, is what the rest of this analysis quantifies.

![Cross-section comparison of sHBM and cHBM](./assets/img-8e41d035.webp)

![Value composition across sHBM, cHBM and NVHBM](./assets/img-857bf4bd.webp)

![The scope of architecture NVIDIA defines](./assets/img-477d9be9.webp)
