---
格式版本: 2
标题: "As AI Grows More Complex, Model Builders Rely on NVIDIA | NVIDIA Blog"
原文链接: "https://blogs.nvidia.com/blog/leading-models-nvidia/"
发布日期: "2026-06-30"
发布时间校准状态: "found"
发布时间来源: "llm:strict_original_body"
发布时间证据: "time class=related-news-date nvidia-article-date datetime=2026-06-30T08:00:57-07:00: Jun 30, 2026"
发布时间校准原因: "该日期位于文章标题下方的元数据区域，且带有 nvidia-article-date 类名，明确标识为文章发布时间。"
发布时间校准置信度: "100"
发布时间候选数量: 41
发布时间严格候选数量: 14
发布时间原页读取状态: "原页面来自已抓取 HTML"
发布时间未找到原因: "候选日期无效或 LLM 未确认"
发布时间校准时间: "2026-07-20T10:24:54+08:00"
发现时间: "2026-07-20T09:24:00+08:00"
入库时间: "2026-07-20T02:36:08.578Z"
来源平台: "NVIDIA Blog 搜索"
搜索渠道: "source_template"
搜索词: "https://blogs.nvidia.com/?s=Scale-up"
匹配关键词:
  - "Scale-up"
  - "GPU"
  - "NVL72"
相关厂家:
  - "NVIDIA"
  - "Microsoft"
  - "Google"
  - "Oracle"
  - "OpenAI"
相关专家:
  []
内容类型: "网页"
抓取工具: "AgentKey Scrape"
清洗工具: "AgentKey Markdown + LLM 正文裁剪"
原始附件:
  []
AI优质: "否"
AI打分: 68
AI分档: "召回候选"
AI质检状态: "不通过"
AI打分理由: "NVIDIA官方博客，提及GB200/GB300 NVL72及Blackwell架构在OpenAI等厂商的部署，但正文侧重模型能力与生态宣传，缺乏超节点/机柜级架构、供电散热互连等具体技术细节，商业信号偏泛。"
AI质检模型: "qwen3.6-plus"
AI质检时间: "2026-07-20T10:36:08+08:00"
AI主题相关性: 12
AI来源权威性: 14
AI新颖性: 14
AI技术细节: 8
AI商业部署信号: 10
AI完整性: 10
图片摘要:
  - "✓ ./assets/img-83d2f103.jpg | diagram | 展示大规模AI服务器集群布局，体现NVIDIA基础设施支撑复杂模型训练与部署的规模化能力。"
  - "★ ./assets/img-1f6575f1.png | diagram | 展示GPU与Vera CPU协同架构及任务分配流程，体现NVIDIA全栈基础设施中的互连拓扑。"
  - "✗ ./assets/img-3511966c.jpg | photo | 品牌宣传照，仅展示NVIDIA Logo及建筑，无具体技术信息。"
  - "✓ ./assets/img-e47e99dc.png | infographic | 展示NVIDIA推理软件栈概念图，体现其在降低Token成本及提升推理效率方面的技术优势。"
采集批次: "2026年7月20日9点23分34秒"
采集批次ID: "20260720-092334-062"
去重键: "https://blogs.nvidia.com/blog/leading-models-nvidia"
---

Unveiling what it describes as the most capable model series yet for professional knowledge work, OpenAI [launched](https://openai.com/index/introducing-gpt-5-2/) GPT-5.2 in December. The model was trained and deployed on NVIDIA infrastructure, including [NVIDIA Hopper](https://www.nvidia.com/en-us/data-center/technologies/hopper-architecture/) and [GB200 NVL72](https://www.nvidia.com/en-us/data-center/gb200-nvl72/) systems.

[GPT-5.3 Codex](https://openai.com/index/introducing-gpt-5-3-codex/) — the first OpenAI agentic coding model to help build itself — was released in February and trained and served entirely on GB200 NVL72.

GPT-5.2 achieves the top reported score for industry benchmarks like GPQA-Diamond, AIME 2025 and Tau2 Telecom. On leading benchmarks targeting the skills required to develop AGI, like [ARC-AGI-2](https://arcprize.org/leaderboard), GPT-5.2 sets a new bar for state-of-the-art performance.

GPT 5.3-Codex combines the coding performance of GPT‑5.2-Codex and the reasoning capabilities of GPT‑5.2 together in one model, with 25% faster performance. In four benchmarks used to evaluate coding, agentic and real-world capabilities, GPT 5.3-Codex set a new industry highs on SWE-Bench Pro and Terminal-Bench while also displaying strong performance on OSWorld and GDPval benchmarks,.

GPT 5.2 and GPT 5.3-Codex are the latest examples of how leading AI builders train and deploy at scale on NVIDIA’s full-stack AI infrastructure.

## Pretraining: The Bedrock of Intelligence

AI models are getting more capable thanks to three [scaling laws](https://blogs.nvidia.com/blog/ai-scaling-laws/): pretraining, post-training and test-time scaling.

[Reasoning models](https://www.nvidia.com/en-us/glossary/ai-reasoning/), which apply compute during [inference](https://www.nvidia.com/en-us/solutions/ai/inference/) to tackle complex queries, using multiple networks working together, are now everywhere.

But pretraining and post-training remain the bedrock of intelligence. They’re core to making reasoning models smarter and more useful.

And getting there takes scale. Training frontier models from scratch isn’t a small job.

It takes tens of thousands, even hundreds of thousands, of GPUs working together effectively.

That level of scale demands excellence across many dimensions. It requires world-class accelerators, advanced networking across scale-up, scale-out and increasingly scale-across architectures, plus a fully optimized software stack. In short, a purpose-built infrastructure platform built to deliver performance at scale.

Compared with the NVIDIA Hopper architecture, NVIDIA GB200 NVL72 systems delivered [3x faster training performance](https://developer.nvidia.com/blog/nvidia-blackwell-enables-3x-faster-training-and-nearly-2x-training-performance-per-dollar-than-previous-gen-architecture/) on [the largest model tested in the latest MLPerf Training industry benchmarks, and nearly 2x better performance per dollar](https://developer.nvidia.com/blog/nvidia-blackwell-enables-3x-faster-training-and-nearly-2x-training-performance-per-dollar-than-previous-gen-architecture/).

And NVIDIA GB300 NVL72 delivers a [more than 4x speedup](https://www.nvidia.com/en-us/data-center/resources/mlperf-benchmarks/) compared with NVIDIA Hopper.

These performance gains help AI developers shorten development cycles and deploy new models more quickly.

## Proof in the Models Across Every Modality

The majority of today’s leading large language models were [trained](https://www.nvidia.com/en-us/solutions/ai/ai-training/) on NVIDIA platforms.

AI isn’t just about text.

NVIDIA supports AI development across multiple modalities, including speech, image and video generation, as well as emerging areas like biology and robotics.

For example, models like [Evo 2](https://blogs.nvidia.com/blog/evo-2-biomolecular-ai/) decode genetic sequences, OpenFold3 predicts 3D protein structures and Boltz-2 simulates drug interactions, helping researchers identify promising candidates faster.

On the clinical side, NVIDIA Clara synthesis models generate realistic medical images to advance screening and diagnosis without exposing patient data.

Companies like [Runway](https://runwayml.com/) and Inworld train on NVIDIA infrastructure.

Runway last week announced Gen-4.5, a new frontier video generation model that’s the current top-rated video model in the world, according to the Artificial Analysis leaderboard.

Now optimized for NVIDIA Blackwell, Gen-4.5 was developed entirely on NVIDIA GPUs across initial research and development, pre-training, post-training and inference.

Runway also announced GWM-1, a state-of-the-art general world model trained on NVIDIA Blackwell that’s built to simulate reality in real time. It’s interactive, controllable and general-purpose, with applications in video games, education, science, entertainment and robotics.

Benchmarks show why.

MLPerf is the industry-standard benchmark for training performance. In the latest round, [NVIDIA submitted results across all seven MLPerf Training 5.1 benchmarks](https://blogs.nvidia.com/blog/mlperf-training-benchmark-blackwell-ultra/), showing strong performance and versatility. It was the only platform to submit in every category.

NVIDIA’s ability to support diverse AI workloads helps data centers use resources more efficiently.

That’s why AI labs such as Black Forest Labs, Cohere, Mistral, OpenAI, Reflection and Thinking Machines Lab and are all training on the NVIDIA Blackwell platform.

## NVIDIA Blackwell Across Clouds and Data Centers

NVIDIA Blackwell is widely available from leading cloud service providers, neo-clouds and server makers.

And NVIDIA Blackwell Ultra, offering additional compute, memory and architecture improvements, is now rolling out from server makers and cloud service providers.

Major cloud service providers and [NVIDIA Cloud Partners](https://www.nvidia.com/en-us/data-center/gpu-cloud-computing/partners/), including Amazon Web Services, CoreWeave, Google Cloud, Lambda, Microsoft Azure, Nebius, Oracle Cloud Infrastructure and Together AI, to name a few, already offer instances powered by NVIDIA Blackwell, ensuring scalable performance as pretraining scaling continues.

From frontier models to everyday AI, the future is being built on NVIDIA.

*Learn more about the* [*NVIDIA Blackwell platform*](https://www.nvidia.com/en-us/data-center/technologies/blackwell-architecture/)*.*

*Editor’s note: This story was updated on February 6, 2026 with the latest model information from OpenAI and its GPT-5.3 Codex. Check back for subsequent model launches and new data from OpenAI.*

![Why Performance per Watt Is the Ultimate Metric for AI Infrastructure Efficiency](./assets/img-83d2f103.jpg)

![AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters](./assets/img-1f6575f1.png)

![How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost](./assets/img-e47e99dc.png)
