---
格式版本: 2
标题: "AMD ROCm 10: Bringing ROCm.AI’s AI-Native Developer Experiences to AMD Platforms"
原文链接: "https://newsroom.amd.com/news/rocm-10-software-ai-native-developer-experiences/"
发布日期: "2026-08-27"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:configured_publication_date_rule"
发布时间证据: "amd-newsroom-publication-date html:original: <time datetime=\"2026-08-27\""
发布时间校准原因: "信源发布日期识别规则直接确认发布时间"
发布时间校准置信度: "high"
发布时间候选数量: 1
发布时间严格候选数量: 1
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-28T00:58:25+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-28T00:58:03+08:00"
入库时间: "2026-08-27T16:59:22.861Z"
来源平台: "固定入口"
搜索渠道: "fixed_url"
搜索词: "https://newsroom.amd.com/"
匹配关键词:
  - "AI"
  - "GPU"
  - "HBM"
  - "performance"
  - "bandwidth"
相关厂家:
  - "AMD"
  - "OpenAI"
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
图片摘要:
  - "★ ./assets/img-18c1c65d.jpg | diagram | 展示ROCm.AI AI驱动开发平台的整体架构，涵盖Hyperloom、AMD Skills与ROCm CLI等核心组件。"
  - "✓ ./assets/img-1583c95d.jpg | screenshot | 展示AMD Skills功能界面，体现将AMD专家经验集成到开发者工具中的能力。"
  - "✓ ./assets/img-fbebc055.jpg | screenshot | 展示ROCm CLI命令行工具界面，体现开发者通过CLI管理AI工作负载的交互方式。"
AI优质: "否"
AI打分: 54
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文主线是ROCm 10及ROCm.AI开发、推理优化和运维工具，并非机架级AI基础设施。AMD官方一手来源且正文完整。历史证据表明ROCm.AI已于Advancing AI 2026亮相；本文新增ROCm 10正式发布、ROCm.AI GA、Hyperloom对vLLM/SGLang等支持、CLI预览及8卡MI355X测试数据。当前页可作为软件发布的一手来源，但没有新机架拓扑、互连、供电、液冷、客户部署或规模交付，性能创新主要面向模型训练与推理优化，命中“应用与模型效率”硬否决项，最高54分。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-08-28T00:59:44+08:00"
AI主题相关性: 3
AI来源权威性: 15
AI新颖性: 14
AI技术细节: 10
AI商业部署信号: 2
AI完整性: 10
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.14"
AI评分知识库SHA256: "da252a8e80bb6f67d5d3464dce82a2da6f260de290a9266154c37ce20774eb0b"
AI评分知识库检索词: "[\"AMD\",\"https://newsroom.amd.com/\",\"HBM\",\"RAS\",\"GPU\",\"Intel\",\"ROCm\",\"ROCm.AI\",\"AI-Native\",\"assets/img-18c1c65d.jpg\",\"AI-driven\",\"CLI\"]"
AI评分知识库命中: "[{\"id\":\"july-correct-0034\",\"title\":\"AMD Fires Back at Nvidia with Helios AI System, Epyc CPUs\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"AMD\",\"https://newsroom.amd.com/\",\"RAS\",\"GPU\",\"Intel\",\"ROCm\",\"ROCm.AI\",\"AI-Native\",\"CLI\"],\"rank\":-17.85177682033857},{\"id\":\"july-correct-0079\",\"title\":\"AAI 2026: AMD Launches AMD Helios Rackscale Solution for Frontier AI\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"AMD\",\"https://newsroom.amd.com/\",\"HBM\",\"RAS\",\"GPU\",\"Intel\",\"ROCm\"],\"rank\":-16.939025592728665},{\"id\":\"july-correct-0081\",\"title\":\"AAI 2026: 6th Gen AMD EPYC Server CPUs Power the Agentic Data Center\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"AMD\",\"https://newsroom.amd.com/\",\"RAS\",\"GPU\",\"Intel\"],\"rank\":-15.043326199776052},{\"id\":\"july-correct-0078\",\"title\":\"AAI 2026: AMD Launches AMD Instinct MI400 Series GPUs for Frontier AI, HPC\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"AMD\",\"https://newsroom.amd.com/\",\"HBM\",\"RAS\",\"GPU\",\"ROCm\"],\"rank\":-11.236656166119253},{\"id\":\"july-correct-0075\",\"title\":\"From EPYC to Helios, AMD and Meta are Scaling the Future of AI Together\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"AMD\",\"https://newsroom.amd.com/\",\"RAS\",\"GPU\",\"ROCm\"],\"rank\":-11.123001893125453}]"
AI摘要: "AMD 发布 ROCm 10，并正式推出 AI 驱动开发平台 ROCm.AI，整合 Hyperloom、AMD Skills 和 ROCm CLI 三类开发体验。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-08-27T18:59:19.096Z"
采集批次: "2026年8月28日0点58分02秒"
采集批次ID: "20260828-005802-203"
去重键: "https://newsroom.amd.com/news/rocm-10-software-ai-native-developer-experiences"
---

![ROCm AI AI Driven Development Platform](./assets/img-18c1c65d.jpg)

ROCm AI AI Driven Development Platform

**Three things to know:**

- [ROCm.AI](http://rocm.ai/) is now generally available with AMD ROCm™ 10 software, bringing an AI-driven development platform to help developers build, run and optimize AI workloads.
- ROCm.AI combines ROCm Hyperloom, AMD Skills and ROCm CLI, bringing AMD expertise, simpler workflows and agentic optimization into the tools developers already use.
- ROCm 10 advances the broader ROCm software stack, with a more modular ROCm Core SDK and enhancements across libraries, compilers, frameworks, model support, performance and hardware platforms.

Today, AMD released AMD ROCm™ 10, marking 10 years of the AMD software stack and making [ROCm.AI](http://rocm.ai/) generally available for users. First introduced at [Advancing AI 2026](https://newsroom.amd.com/press-kits/advancing-ai-2026-all-news/), ROCm.AI is an AI-native software experience designed to accelerate development velocity and performance optimization on AMD hardware.

ROCm.AI brings together three core developer experiences: AMD Skills, the ROCm CLI and AMD Hyperloom. Together, they introduce AMD expertise and agentic workflows into the tools developers already use, enabling innovation to evolve at the pace of AI.

Through AI-driven optimization of kernels, memory management and scheduling, a system configured with ROCm.AI delivers an average 3.3x inference improvement and 2.4x training improvement over ROCm™ 7 on the same hardware.[^1] [^2]

## A New Developer Experience for AI on AMD

ROCm.AI supports developers across three parts of the AI development process: optimizing inference performance, building with AMD expertise and running AI workloads.

**Optimize End-to-End Inference with AMD ROCm Hyperloom**

ROCm Hyperloom, a key component of ROCm.AI, is an autonomous agentic system for optimizing end-to-end inference workloads across host code and GPU kernels.

Hyperloom profiles workloads, identifies bottlenecks, explores optimization options, implements targeted changes, benchmarks the results, and validates performance and correctness.

AMD Hyperloom

With ROCm 10, Hyperloom expands support across AMD Instinct™ GPUs, with support for vLLM and SGLang. Developers can target optimizations across HIP, Triton and FlyDSL, with reports describing proposed code changes and measured or expected performance improvements.

Hyperloom is available through standalone workflows and AMD Skills, giving developers multiple ways to incorporate agentic optimization into their development process.

**Build with AMD Expertise**

AMD Skills brings curated AMD knowledge and validated workflows into supported AI coding agents like Claude Code, Cursor and Codex, giving developers AMD-specific guidance within the tools they already use. With ROCm 10, the AMD Skills catalog expands across three areas:

- **Client-native workflows** for local AI and application integration.
- **Cross-stack workflows** for diagnostics, routing, replay analysis and optimization.
- **Server-native workflows** for AMD Instinct GPUs and AMD EPYC™ processors, including serving, profiling and performance analysis.
![AMD Skills](./assets/img-1583c95d.jpg)

AMD Skills

The skills previewed at Advancing AI are now available through the Claude Code, Codex and Cursor marketplaces, as well as through the open catalog on [GitHub](https://github.com/amd/skills). Each shipped skill passes structural and behavioral testing before release to help provide consistent, reliable workflows.

**Run with Simpler Workflows**

The ROCm CLI, a Technology Preview component of ROCm.AI, provides a stable, unified command-line interface for setting up, managing and operating AI workloads on AMD hardware.

Developers can use the same workflows manually, through an AI coding agent or in continuous integration (CI) environments. From a single interface, they can inspect systems, install and manage ROCm environments, serve models, run diagnostics, update components, and control runtimes.

![AMD ROCm CLI](./assets/img-fbebc055.jpg)

AMD ROCm CLI

The CLI is available on Windows as well as Linux as a prebuilt binary and does not require an existing ROCm installation. It delivers managed ROCm environments, supports multiple side-by-side runtimes including runtime activation and rollback, as well as integrated model serving and engine management. Adapters today support Lemonade on select AMD client systems and vLLM for AMD Instinct GPU serving.

The ROCm Console (formerly “dash”), included with the CLI, provides a real-time view of system status and workload activity. Developers can monitor ROCm runtime health, model serving, GPU utilization and benchmark telemetry, including metrics such as high bandwidth memory (HBM) usage, power consumption and tokens per watt on supported AMD Instinct systems.

Available as a Technology Preview, ROCm CLI provides a version-agnostic experience across ROCm releases, beginning with ROCm 7.13 software, with ROCm 10 official support coming soon.

## Building on the ROCm 10 Software Stack

These developer experiences build on the broader ROCm 10 software stack that includes the ROCm Core SDK, providing a more modular foundation for building and running AI workloads across AMD platforms.

ROCm 10 also includes updates to libraries, compilers, frameworks, tools, model support, performance and hardware platforms.

For a deeper look at the ROCm Core SDK, including libraries, compilers, tools, framework and model support, performance improvements and platform updates, read the full [ROCm 10 technical deep dive](https://advanced-micro-devices-rocm-blogs--118.com.readthedocs.build/projects/preview/en/118/ecosystems-and-partners/rocm-x-blog/README.html).

Topics [Software](https://newsroom.amd.com/category/software/ "Primary category: Software") [Artificial Intelligence](https://newsroom.amd.com/category/ai/ "Also filed under: Artificial Intelligence") [#rocm](https://newsroom.amd.com/tag/rocm/) [#epyc](https://newsroom.amd.com/tag/epyc/) [#instinct](https://newsroom.amd.com/tag/instinct/)

Press inquiries: [corporate.pressinquiry@amd.com](mailto:corporate.pressinquiry@amd.com)

[^1]: (MI350-81) Testing by AMD Performance Labs as of July 7, 2026, measuring the inference performance in tokens per second (TPS) of a system configured with an AMD Instinct MI355x 8x GPU platform and AMD ROCm 7.0 software vs a similarly configured system using a preview version of AMD [ROCm.ai](http://rocm.ai/) (ROCm 7.2.2 with optimizations such as Optimized Kernels, Parallelism and Scheduling) running GLM-5, Kimi-K2.5, and DeepSeekk-R1-0528 models.

Stated performance uplift is expressed as a combined average TPS over across the (3) models tested.

Hardware Configuration

Supermicro AS -4126GS-NMR-LCC (board H14DSG-OD)

8x AMD Instinct MI355X. BIOS AMI v1.4a (2025-04-16), GPU firmware SMC 04.86.11.02, TA RAS 27.69.00.10, TA XGMI 32.00.00.20, RLC43, MEC36, SDMA12, Ubuntu 22.04.2 LTS, kernel 5.15.0-70-generic, amdgpu driver 6.16.6, HOST ROCm 7.1.0

Software Configuration(s)

GLM-5 ROCm Docker Image: rocm/sgl-dev:v0.5.8.post1-rocm700-mi35x-20260219

PYTorch Version 2.8.0, SGLang v0.5.8

Kimi ROCm Docker Image: vllm/vllm-openai-rocm:v0.16.0, vLLM version 0.16.0

DeepSeek-R1 Docker Image: rocm/7.0:...sgl-dev-v0.5.2-rocm7.0-mi35x-20250915, SGLang version V0.5.13

vs

GLM-5 ROCm Docker Image: rocm/atom:rocm7.2.2\_ubuntu24.04\_py3.12\_pytorch\_release\_2.10.0\_ [atom0.1.2.post](http://atom0.1.2.post/), ATOM [v0.1.2.post](http://v0.1.2.post/)

Kimi ROCm Docker Image: vllm/vllm-openai-rocm:v0.22.0, vLLM version V0.22.0

DeepSeek-R1 Docker Image: lmsysorg/sglang-rocm:v0.5.13-rocm720-mi35x-20260612, SGLang version V0.5.13

Server manufacturers may vary configurations, yielding different results. Performance may vary based on configuration, software, vLLM version, and the use of the latest drivers and optimizations. (MI350-81)

[^2]: (MI350-82) Testing by AMD Performance Labs as of July 7, 2026, measuring the training performance in tokens per second (TPS) of AMD ROCm 7.0 software vs a preview version of AMD [ROCm.ai](http://rocm.ai/) (ROCm 7.2.2 with optimizations such as Optimized Kernels, Parallelism and Scheduling),) using Megatron -LM on a system with 8x AMD Instinct MI355x 8x GPUs GPU platform running DeepSeek-V2-Lite, DeepSeek-V3-16B, and Qwen3-30B-A3B models.

Stated performance uplift is expressed as the combined average TPS over across the (3) models tested.

Hardware Configuration

Supermicro AS -4126GS-NMR-LCC (board H14DSG-OD)

8x AMD Instinct MI355. BIOS AMI v1.4a (2025-04-16), GPU firmware SMC 04.86.11.02, TA RAS 27.69.00.10, TA XGMI 32.00.00.20, RLC 43, MEC 36, SDMA 12, Ubuntu 22.04.2 LTS, kernel 5.15.0-70-generic, amdgpu driver 6.16.6, HOST ROCm 7.1.0

Software Configuration(s)

DeepSeek-V2-Lite, ROCm 7.2.1 + Primus v26.3

DeepSeek-V3-16B, ROCm 7.2.1 + Primus v26.3

Qwen3-30B-A3B, ROCm 7.2.1 + Primus v26.3

Server manufacturers may vary configurations, yielding different results. Performance may vary based on configuration, software, and the use of the latest drivers and optimizations. (MI350-82)

![ROCm AI AI Driven Development Platform](./assets/img-18c1c65d.jpg)

![AMD Skills](./assets/img-1583c95d.jpg)

![AMD ROCm CLI](./assets/img-fbebc055.jpg)
