---
格式版本: 2
标题: "Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration"
原文链接: "https://arxiv.org/abs/2608.24664"
发布日期: "2026-08-25"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "[Submitted on 25 Aug 2026]"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 8
发布时间严格候选数量: 2
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-26T21:55:17+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-26T21:50:45+08:00"
入库时间: "2026-08-26T13:55:30.500Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=HBM&searchtype=all"
匹配关键词:
  - "HBM"
  - "performance"
  - "bandwidth"
  - "AI"
相关厂家:
  - "Google"
相关专家:
  []
内容类型: "网页"
抓取工具: "Jina Reader"
清洗工具: "Jina Reader Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
图片摘要:
  - "✗ ./assets/img-5838d115.png | ad | 与正文无关的基金会品牌标识"
  - "✗ ./assets/img-26707af3.png | ad | 与正文无关的基金会品牌标识"
  - "✗ ./assets/img-43a00bd2.png | ad | 与正文无关的基金会品牌标识"
AI优质: "否"
AI打分: 66
AI分档: "召回候选"
AI质检状态: "不通过"
AI打分理由: "正文主线是新型Maia 200 AI加速器及SDLA数据流架构，属于机架级AI基础设施的关键计算部件，但未直接展开机架拓扑或集群系统。arXiv原始论文页面来源可追溯，但当前仅提供摘要。相较固定知识库，新增Maia 200的10145 TFLOP/s FP4、5072 TFLOP/s FP8、750W TDP、7 TB/s HBM带宽及软件定义数据流机制，未发现重复事件；但没有客户、量产、交付、真实部署或端到端实测证据，当前页面不足以命中高置信准入通道。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-08-26T21:56:05+08:00"
AI主题相关性: 14
AI来源权威性: 13
AI新颖性: 18
AI技术细节: 13
AI商业部署信号: 1
AI完整性: 7
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.3"
AI评分知识库SHA256: "dbc02c7552b478ae5aae514533e58a5de9ce72d2911e4aaac0f744c08178e4a6"
AI评分知识库检索词: "[\"HBM\",\"RAS\",\"Google\",\"Intel\",\"URL\",\"arxiv.org/abs/2608.24664\",\"HTML\",\"PDF\",\"arxiv.org/pdf/2608.24664\",\"arxiv.org/html/2608.24664v1\",\"performance-10\",\"FP4\"]"
AI评分知识库命中: "[{\"id\":\"historical-jan-apr-02\",\"title\":\"二、Google Cloud Next '26：AI Hypercomputer 与第八代 TPU 发布\",\"sourceType\":\"curated_item\",\"time\":\"2026-01_to_2026-04\",\"matchedTerms\":[\"Google\",\"Intel\"],\"rank\":-11.38690686659772},{\"id\":\"july-correct-0081\",\"title\":\"AAI 2026: 6th Gen AMD EPYC Server CPUs Power the Agentic Data Center\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"RAS\",\"Intel\",\"URL\",\"HTML\",\"PDF\"],\"rank\":-10.81660765764746},{\"id\":\"july-correct-0111\",\"title\":\"StrataCL: Fabric-Native Communication Library for Production Supernodes\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"URL\",\"HTML\",\"PDF\"],\"rank\":-8.0985860723721},{\"id\":\"july-correct-0026\",\"title\":\"The Rackscale AI System Roadmaps That AMD Is Using To Chase Money\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"HBM\",\"RAS\",\"Intel\",\"URL\",\"FP4\"],\"rank\":-6.921015500961769},{\"id\":\"july-correct-0079\",\"title\":\"AAI 2026: AMD Launches AMD Helios Rackscale Solution for Frontier AI\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"HBM\",\"RAS\",\"Intel\",\"URL\",\"HTML\",\"FP4\"],\"rank\":-6.796760126111113}]"
AI摘要: "Maia 200 论文介绍了一款面向大规模 AI 推理的加速器，采用软件定义数据流架构，将设计从线程为中心转向数据移动为中心，以提高效率与可扩展性。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-08-27T08:12:41.374Z"
采集批次: "2026年8月26日20点37分07秒"
采集批次ID: "20260826-203707-934"
去重键: "https://arxiv.org/abs/2608.24664"
---

Title: Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration

URL Source: https://arxiv.org/abs/2608.24664

Markdown Content:
[Skip to main content](https://arxiv.org/abs/2608.24664#content)[](https://arxiv.org/IgnoreMe)[Search](https://arxiv.org/search)[Submit](https://arxiv.org/user/create)[Donate](https://info.arxiv.org/about/donate.html)[Log in](https://arxiv.org/login)

Search arXiv 

 Press Enter to search · [Advanced search](https://arxiv.org/search/advanced)

# Computer Science > Hardware Architecture

**arXiv:2608.24664** (cs) 

 [Submitted on 25 Aug 2026]

# Title:Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration

Authors:[Sherry Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+S), [Marco Heddes](https://arxiv.org/search/cs?searchtype=author&query=Heddes,+M), [Jackson Peng](https://arxiv.org/search/cs?searchtype=author&query=Peng,+J), [Tom Savell](https://arxiv.org/search/cs?searchtype=author&query=Savell,+T), [Monica Tang](https://arxiv.org/search/cs?searchtype=author&query=Tang,+M), [Prashant Ranjan](https://arxiv.org/search/cs?searchtype=author&query=Ranjan,+P), [Jesse Benson](https://arxiv.org/search/cs?searchtype=author&query=Benson,+J), [Ofer Dekel](https://arxiv.org/search/cs?searchtype=author&query=Dekel,+O), [Saurabh Dighe](https://arxiv.org/search/cs?searchtype=author&query=Dighe,+S), [Anupama Kurpad](https://arxiv.org/search/cs?searchtype=author&query=Kurpad,+A), [Artour Levin](https://arxiv.org/search/cs?searchtype=author&query=Levin,+A), [Matthew Mattina](https://arxiv.org/search/cs?searchtype=author&query=Mattina,+M), [George Petre](https://arxiv.org/search/cs?searchtype=author&query=Petre,+G), [Cheng Tang](https://arxiv.org/search/cs?searchtype=author&query=Tang,+C), [Yuan Yu](https://arxiv.org/search/cs?searchtype=author&query=Yu,+Y), [Li Zhang](https://arxiv.org/search/cs?searchtype=author&query=Zhang,+L), [Torsten Hoefler](https://arxiv.org/search/cs?searchtype=author&query=Hoefler,+T)

View a PDF of the paper titled Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration, by Sherry Xu and 16 other authors

[View PDF](https://arxiv.org/pdf/2608.24664)[HTML (experimental)](https://arxiv.org/html/2608.24664v1)
> Abstract:We introduce Maia 200, an advanced AI accelerator delivering high performance-10 145 Tflop/s FP4 and 5072 Tflop/s FP8 within a 750W TDP and 7 TB/s HBM bandwidth. Maia exemplifies a new class of Software Defined Locally Accessed Dataflow Architectures (SDLA), which explicitly program dataflow engines to orchestrate highly specialized memories and data movement engines. This approach shifts the focus from today's thread-centric to data-movement-centric architecture, improving efficiency and scalability. Our taxonomy of data management, inspired by Flynn's classification, highlights how SDLA addresses challenges in modern AI computing. Maia 200 achieves significant cost and energy savings while supporting massive parallelism for AI inference workloads, making it a compelling solution for next-generation high-performance computing systems.

Subjects:Hardware Architecture (cs.AR); Artificial Intelligence (cs.AI); Distributed, Parallel, and Cluster Computing (cs.DC); Emerging Technologies (cs.ET); Machine Learning (cs.LG)
Cite as:[arXiv:2608.24664](https://arxiv.org/abs/2608.24664) [cs.AR]
(or [arXiv:2608.24664v1](https://arxiv.org/abs/2608.24664v1) [cs.AR] for this version)
[https://doi.org/10.48550/arXiv.2608.24664](https://doi.org/10.48550/arXiv.2608.24664)

Focus to learn more

 arXiv-issued DOI via DataCite (pending registration)

## Submission history

 From: Torsten Hoefler [[view email](https://arxiv.org/show-email/8e24be81/2608.24664)] 

**[v1]** Tue, 25 Aug 2026 15:05:40 UTC (1,145 KB)

[](https://arxiv.org/abs/2608.24664)Full-text links:
## Access Paper:

 View a PDF of the paper titled Maia 200: A Software Defined Dataflow System for Large-scale AI Acceleration, by Sherry Xu and 16 other authors

*   [View PDF](https://arxiv.org/pdf/2608.24664)
*   [HTML (experimental)](https://arxiv.org/html/2608.24664v1)
*   [TeX Source](https://arxiv.org/src/2608.24664)

[view license](http://creativecommons.org/licenses/by/4.0/ "Rights to this article")

### Current browse context:

cs.AR

[<prev](https://arxiv.org/prevnext?id=2608.24664&function=prev&context=cs.AR "previous in cs.AR (accesskey p)") | [next>](https://arxiv.org/prevnext?id=2608.24664&function=next&context=cs.AR "next in cs.AR (accesskey n)")

[new](https://arxiv.org/list/cs.AR/new) | [recent](https://arxiv.org/list/cs.AR/recent) | [2026-08](https://arxiv.org/list/cs.AR/2026-08)

 Change to browse by: 

[cs](https://arxiv.org/abs/2608.24664?context=cs)

[cs.AI](https://arxiv.org/abs/2608.24664?context=cs.AI)

[cs.DC](https://arxiv.org/abs/2608.24664?context=cs.DC)

[cs.ET](https://arxiv.org/abs/2608.24664?context=cs.ET)

[cs.LG](https://arxiv.org/abs/2608.24664?context=cs.LG)

### References & Citations

*   [NASA ADS](https://ui.adsabs.harvard.edu/abs/arXiv:2608.24664)
*   [Google Scholar](https://scholar.google.com/scholar_lookup?arxiv_id=2608.24664)
*   [Semantic Scholar](https://api.semanticscholar.org/arXiv:2608.24664)

export BibTeX citation Loading...

## BibTeX formatted citation

×

Data provided by: [](https://arxiv.org/abs/2608.24664)

### Bookmark

Bibliographic Tools 

# Bibliographic and Citation Tools

- [x] Bibliographic Explorer Toggle 

Bibliographic Explorer _([What is the Explorer?](https://info.arxiv.org/labs/showcase.html#arxiv-bibliographic-explorer))_

- [x] Connected Papers Toggle 

Connected Papers _([What is Connected Papers?](https://www.connectedpapers.com/about))_

- [x] Litmaps Toggle 

Litmaps _([What is Litmaps?](https://www.litmaps.co/))_

- [x] scite.ai Toggle 

scite Smart Citations _([What are Smart Citations?](https://www.scite.ai/))_

Code, Data, Media 

# Code, Data and Media Associated with this Article

- [x] alphaXiv Toggle 

alphaXiv _([What is alphaXiv?](https://alphaxiv.org/))_

- [x] Links to Code Toggle 

CatalyzeX Code Finder for Papers _([What is CatalyzeX?](https://www.catalyzex.com/))_

- [x] DagsHub Toggle 

DagsHub _([What is DagsHub?](https://dagshub.com/))_

- [x] GotitPub Toggle 

Gotit.pub _([What is GotitPub?](http://gotit.pub/faq))_

- [x] Huggingface Toggle 

Hugging Face _([What is Huggingface?](https://huggingface.co/huggingface))_

- [x] ScienceCast Toggle 

ScienceCast _([What is ScienceCast?](https://sciencecast.org/welcome))_

Demos 

# Demos

- [x] Replicate Toggle 

Replicate _([What is Replicate?](https://replicate.com/docs/arxiv/about))_

- [x] Spaces Toggle 

Hugging Face Spaces _([What is Spaces?](https://huggingface.co/docs/hub/spaces))_

- [x] Spaces Toggle 

TXYZ.AI _([What is TXYZ.AI?](https://txyz.ai/))_

Related Papers 

# Recommenders and Search Tools

- [x] Link to Influence Flower 

Influence Flower _([What are Influence Flowers?](https://influencemap.cmlab.dev/))_

- [x] Core recommender toggle 

CORE Recommender _([What is CORE?](https://core.ac.uk/services/recommender))_

*   [Author](https://arxiv.org/abs/2608.24664)
*   [Venue](https://arxiv.org/abs/2608.24664)
*   [Institution](https://arxiv.org/abs/2608.24664)
*   [Topic](https://arxiv.org/abs/2608.24664)

 About arXivLabs  

# arXivLabs: experimental projects with community collaborators

arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community? [**Learn more about arXivLabs**](https://info.arxiv.org/labs/index.html).

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.24664) | [Disable MathJax](javascript:setMathjaxCookie()) ([What is MathJax?](https://info.arxiv.org/help/mathjax.html)) 

 We gratefully acknowledge support from our **major funders**, [**member institutions**](https://info.arxiv.org/about/ourmembers.html), , and all contributors. 

[About](https://info.arxiv.org/about)·[Help](https://info.arxiv.org/help)·[Contact](https://info.arxiv.org/help/contact.html)·[Subscribe](https://info.arxiv.org/help/subscribe)·[Copyright](https://info.arxiv.org/help/license/index.html)·[Privacy](https://info.arxiv.org/help/policies/privacy_policy.html)·[Accessibility](https://info.arxiv.org/help/web_accessibility.html)·[Operational Status (opens in new tab)](https://status.arxiv.org/)

Major funding support from
