---
格式版本: 2
标题: "OCI Expands NVIDIA GPU-Accelerated Instances for AI, Digital Twins | NVIDIA Blog"
原文链接: "https://blogs.nvidia.com/blog/oracle-cloud-infrastructure-ai-gpu-digital-twins/"
发布日期: "2026-07-14"
发布时间校准状态: "found"
发布时间来源: "llm:original_extracted_text"
发布时间证据: "第 1 行：original_dom_semantic: time class=related-news-date nvidia-article-date datetime=2026-07-14T08:00:52-07:00: Jul 14, 2026"
发布时间校准原因: "候选日期位于标题下方/元数据区域，且为最新发布日期，符合文章发布时间特征。"
发布时间校准置信度: "1"
发布时间候选数量: 36
发布时间严格候选数量: 12
发布时间原页读取状态: "原页面来自已抓取 HTML"
发布时间未找到原因: "候选日期无效或 LLM 未确认"
发布时间校准时间: "2026-07-20T13:24:23+08:00"
发现时间: "2026-07-20T09:26:32+08:00"
入库时间: "2026-07-20T05:29:58.402Z"
来源平台: "NVIDIA Blog 搜索"
搜索渠道: "source_template"
搜索词: "https://blogs.nvidia.com/?s=Oracle"
匹配关键词:
  - "GPU"
  - "Nvlink"
相关厂家:
  - "Oracle"
  - "NVIDIA"
相关专家:
  []
内容类型: "网页"
抓取工具: "AgentKey Scrape"
清洗工具: "AgentKey Markdown + LLM 正文裁剪"
原始附件:
  []
AI优质: "否"
AI打分: 48
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "文章主要介绍OCI云上新增的L40S裸金属实例、单卡H100虚拟机及GH200测试实例，属于常规云实例发布，未涉及超节点、AI Rack、机柜级系统架构、供电散热或高速互连等核心内容，技术细节偏应用层，商业信号为常规云服务上线，与项目关注…"
AI质检模型: "qwen3.6-plus"
AI质检时间: "2026-07-20T13:29:58+08:00"
AI主题相关性: 8
AI来源权威性: 14
AI新颖性: 10
AI技术细节: 6
AI商业部署信号: 6
AI完整性: 4
图片摘要:
  - "✗ ./assets/img-83d2f103.jpg | diagram | 与正文主题无关，Alt text指向能效话题，疑似推荐阅读配图"
  - "✗ ./assets/img-1f6575f1.png | diagram | 与正文主题无关，Alt text指向Vera CPU，疑似推荐阅读配图"
  - "✗ ./assets/img-3511966c.jpg | photo | 与正文主题无关，品牌宣传图，疑似推荐阅读配图"
  - "✗ ./assets/img-e47e99dc.png | infographic | 与正文主题无关，Alt text指向推理成本，疑似推荐阅读配图"
采集批次: "2026年7月20日9点23分34秒"
采集批次ID: "20260720-092334-062"
去重键: "https://blogs.nvidia.com/blog/oracle-cloud-infrastructure-ai-gpu-digital-twins"
---

Enterprises are rapidly adopting [generative AI](https://www.nvidia.com/en-us/glossary/generative-ai/), large language models ([LLMs](https://www.nvidia.com/en-us/glossary/large-language-models/)), advanced graphics and [digital twins](https://www.nvidia.com/en-us/glossary/digital-twin/) to increase operational efficiencies, reduce costs and drive innovation.

However, to adopt these technologies effectively, enterprises need access to state-of-the-art, full-stack accelerated computing platforms. To meet this demand, Oracle Cloud Infrastructure (OCI) today announced [NVIDIA L40S GPU](https://www.nvidia.com/en-us/data-center/l40s/) bare-metal instances available to order and the upcoming availability of a new virtual machine accelerated by a single [NVIDIA H100 Tensor Core GPU](https://www.nvidia.com/en-us/data-center/h100/). This new VM expands OCI’s existing H100 portfolio, which includes an NVIDIA HGX H100 8-GPU bare-metal instance.

Paired with NVIDIA networking and running the NVIDIA software stack, these platforms deliver powerful performance and efficiency, enabling enterprises to advance generative AI.

## NVIDIA L40S Now Available to Order on OCI

The NVIDIA L40S is a universal data center GPU designed to deliver breakthrough multi-workload acceleration for generative AI, graphics and video applications. Equipped with fourth-generation Tensor Cores and support for the FP8 data format, the L40S GPU excels in training and fine-tuning small- to mid-size LLMs and in inference across a wide range of generative AI use cases.

For example, a single L40S GPU (FP8) can generate up to [1.4x more tokens per second](https://github.com/NVIDIA/TensorRT-LLM/blob/main/docs/source/performance/perf-overview.md) than a single [NVIDIA A100 Tensor Core GPU](https://www.nvidia.com/en-us/data-center/a100/) (FP16) for Llama 3 8B with [NVIDIA TensorRT-LLM](https://developer.nvidia.com/tensorrt) at an input and output sequence length of 128.

The L40S GPU also has best-in-class graphics and media acceleration. Its third-generation NVIDIA Ray Tracing Cores (RT Cores) and multiple encode/decode engines make it ideal for advanced visualization and digital twin applications.

The L40S GPU delivers up to 3.8x the real-time ray-tracing performance of its predecessor, and supports NVIDIA DLSS 3 for faster rendering and smoother frame rates. This makes the GPU ideal for developing applications on the [NVIDIA Omniverse](https://www.nvidia.com/en-us/omniverse/) platform, enabling real-time, photorealistic 3D simulations and AI-enabled digital twins. With Omniverse on the L40S GPU, enterprises can develop advanced 3D applications and workflows for industrial digitalization that will allow them to design, simulate and optimize products, processes and facilities in real time before going into production.

OCI will offer the L40S GPU in its BM.GPU.L40S.4 bare-metal compute shape, featuring four NVIDIA L40S GPUs, each with 48GB of GDDR6 memory. This shape includes local NVMe drives with 7.38TB capacity, 4th Generation Intel Xeon CPUs with 112 cores and 1TB of system memory.

These shapes eliminate the overhead of any virtualization for high-throughput and latency-sensitive AI or machine learning workloads with OCI’s bare-metal compute architecture. The accelerated compute shape features the [NVIDIA BlueField-3 DPU](https://www.nvidia.com/en-us/networking/products/data-processing-unit/) for improved server efficiency, offloading data center tasks from CPUs to accelerate networking, storage and security workloads. The use of BlueField-3 DPUs furthers OCI’s strategy of off-box virtualization across its entire fleet.

[OCI Supercluster](https://www.oracle.com/ai-infrastructure/) with NVIDIA L40S enables ultra-high performance with 800Gbps of internode bandwidth and low latency for up to 3,840 GPUs. OCI’s cluster network uses [NVIDIA ConnectX-7 NICs](https://www.nvidia.com/en-us/networking/ethernet-adapters/) over RoCE v2 to support high-throughput and latency-sensitive workloads, including AI training.

“We chose OCI AI infrastructure with bare-metal instances and NVIDIA L40S GPUs for 30% more efficient video encoding,” said Sharon Carmel, CEO of Beamr Cloud. “Videos processed with Beamr Cloud on OCI will have up to 50% reduced storage and network bandwidth consumption, speeding up file transfers by 2x and increasing productivity for end users. Beamr will provide OCI customers video AI workflows, preparing them for the future of video.”

## Single-GPU H100 VMs Coming Soon on OCI

The VM.GPU.H100.1 compute virtual machine shape, accelerated by a single [NVIDIA H100 Tensor Core GPU](https://www.nvidia.com/en-us/data-center/h100/), is coming soon to OCI. This will provide cost-effective, on-demand access for enterprises looking to use the power of NVIDIA H100 GPUs for their generative AI and HPC workloads.

A single H100 provides a good platform for smaller workloads and LLM inference. For example, one H100 GPU can generate more than [27,000 tokens per second](https://github.com/NVIDIA/TensorRT-LLM/blob/main/docs/source/performance/perf-overview.md) for Llama 3 8B (up to 4x more throughput than a single A100 GPU at FP16 precision) with [NVIDIA TensorRT-LLM](https://developer.nvidia.com/tensorrt) at an input and output sequence length of 128 and FP8 precision.

The VM.GPU.H100.1 shape includes 2×3.4TB of NVMe drive capacity, 13 cores of 4th Gen Intel Xeon processors and 246GB of system memory, making it well-suited for a range of AI tasks.

“Oracle Cloud’s bare-metal compute with NVIDIA H100 and A100 GPUs, low-latency Supercluster and high-performance storage delivers up to 20% better price-performance for Altair’s computational fluid dynamics and structural mechanics solvers,” said Yeshwant Mummaneni, chief engineer of data management analytics at Altair. “We look forward to leveraging these GPUs with virtual machines for the Altair Unlimited virtual appliance.”

## GH200 Bare-Metal Instances Available for Validation

OCI has also made available the BM.GPU.GH200 compute shape for customer testing. It features the [NVIDIA Grace Hopper Superchip](https://www.nvidia.com/en-us/data-center/grace-hopper-superchip/) and [NVLink-C2C](https://www.nvidia.com/en-us/data-center/nvlink-c2c/), a high-bandwidth, cache-coherent 900GB/s connection between the [NVIDIA Grace CPU](https://www.nvidia.com/en-us/data-center/grace-cpu/) and [NVIDIA Hopper](https://www.nvidia.com/en-us/data-center/technologies/hopper-architecture/) GPU. This provides over 600GB of accessible memory, enabling up to 10x higher performance for applications running terabytes of data compared to the NVIDIA A100 GPU.

## Optimized Software for Enterprise AI

Enterprises have a wide variety of NVIDIA GPUs to accelerate their AI, HPC and data analytics workloads on OCI. However, maximizing the full potential of these GPU-accelerated compute instances requires an optimized software layer.

[NVIDIA NIM](https://developer.nvidia.com/blog/nvidia-nim-offers-optimized-inference-microservices-for-deploying-ai-models-at-scale/), part of the [NVIDIA AI Enterprise software platform available on the OCI Marketplace](https://cloudmarketplace.oracle.com/marketplace/en_US/listing/155314141), is a set of easy-to-use microservices designed for secure, reliable deployment of high-performance AI model inference to deploy world-class generative AI applications.

Optimized for NVIDIA GPUs, NIM pre-built containers offer developers improved cost of ownership, faster time to market and security. NIM microservices for popular community models, found on the [NVIDIA API Catalog](https://build.nvidia.com/explore/discover), can be deployed easily on OCI.

Performance will continue to improve over time with upcoming GPU-accelerated instances, including NVIDIA H200 Tensor Core GPUs and NVIDIA Blackwell GPUs.

Order the L40S GPU and test the GH200 Superchip by [reaching out to OCI](https://www.oracle.com/cloud/contact-form.html). To learn more, join [Oracle and NVIDIA at SIGGRAPH](https://go.oracle.com/LP=143193), the world’s premier graphics conference, running through Aug. 1.

*See* [*notice*](https://nam11.safelinks.protection.outlook.com/?url=https%3A%2F%2Fwww.nvidia.com%2Fen-us%2Fabout-nvidia%2Flegal-info%2F&data=05%7C02%7Clpham%40nvidia.com%7Cd59f2f66f51e4deaac8008dc94b3ef0f%7C43083d15727340c1b7db39efd9ccc17a%7C0%7C0%7C638548747745016311%7CUnknown%7CTWFpbGZsb3d8eyJWIjoiMC4wLjAwMDAiLCJQIjoiV2luMzIiLCJBTiI6Ik1haWwiLCJXVCI6Mn0%3D%7C0%7C%7C%7C&sdata=dm8Os%2B4LtHW2ehZrPaxn38bsutMQBDeUdQuxrIa2y1Y%3D&reserved=0) *regarding software product information.*
