---
格式版本: 2
标题: "SatDL: Jointly Optimizing Data Redistribution and Training for Satellite-Based Distributed Learning"
原文链接: "https://arxiv.org/abs/2608.24516"
发布日期: "2026-08-25"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Tue, 25 Aug 2026 13:06:24 UTC (2,767 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-26T21:53:33+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-26T21:50:39+08:00"
入库时间: "2026-08-26T13:53:33.164Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=Scale-up&searchtype=all"
匹配关键词:
  - "Scale-up"
相关厂家:
  - "NVIDIA"
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 34
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文主线是卫星端分布式学习的数据重分配与训练联合优化，而非超节点、AI Rack或机架级基础设施。来源为作者提交的arXiv论文页面，具备可追溯性但尚非正式同行评审成果；提出SatDL及Distributor-Critic框架，并以1584颗卫星仿真和Jetson/A100硬件模拟报告时延、能耗改善，固定知识库未见同项成果，但不能据此认定首次出现。当前仅有摘要，且没有生产部署、客户、量产或机架级架构信息，命中学术模拟及单一学习任务优化强否决项。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-08-26T21:53:56+08:00"
AI主题相关性: 1
AI来源权威性: 10
AI新颖性: 12
AI技术细节: 5
AI商业部署信号: 0
AI完整性: 6
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.3"
AI评分知识库SHA256: "dbc02c7552b478ae5aae514533e58a5de9ce72d2911e4aaac0f744c08178e4a6"
AI评分知识库检索词: "[\"Scale-up\",\"GPU\",\"NVIDIA\",\"PDF\",\"arxiv.org/pdf/2608.24516\",\"HTML\",\"arxiv.org/html/2608.24516v1\",\"IID\",\"A100\",\"GPUs\",\"arxiv.org/abs/2608.24516\",\"arxiv.org/abs/2608.24516v1\"]"
AI评分知识库命中: "[{\"id\":\"july-correct-0089\",\"title\":\"Setting a World Record for MoE Pre-Training on NVIDIA GB300 NVL72\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"GPU\",\"NVIDIA\",\"HTML\",\"GPUs\"],\"rank\":-9.360241728473463},{\"id\":\"july-correct-0081\",\"title\":\"AAI 2026: 6th Gen AMD EPYC Server CPUs Power the Agentic Data Center\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"GPU\",\"NVIDIA\",\"PDF\",\"HTML\",\"GPUs\"],\"rank\":-9.311990673880034},{\"id\":\"july-correct-0027\",\"title\":\"Microsoft Taps AMD For At Scale AI CPU And GPU Clusters\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"GPU\",\"NVIDIA\",\"GPUs\"],\"rank\":-7.887088848905212},{\"id\":\"july-correct-0015\",\"title\":\"AMD to join the optical interconnect party with 2027 Instinct GPUs\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"GPU\",\"NVIDIA\",\"GPUs\"],\"rank\":-7.326474204041515},{\"id\":\"july-correct-0095\",\"title\":\"Inside NVIDIA Rubin GPU Architecture: Powering the Era of Agentic AI\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"Scale-up\",\"GPU\",\"NVIDIA\",\"GPUs\"],\"rank\":-7.305732581415249}]"
AI摘要: "论文提出并评估卫星分布式学习框架SatDL，用Distributor-Critic机制联合优化数据重分发与训练时间，克服星上标签不均衡导致的训练慢、能耗高问题。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-08-27T08:12:55.732Z"
采集批次: "2026年8月26日20点37分07秒"
采集批次ID: "20260826-203707-934"
去重键: "https://arxiv.org/abs/2608.24516"
---

## Computer Science > Distributed, Parallel, and Cluster Computing

## Title:SatDL: Jointly Optimizing Data Redistribution and Training for Satellite-Based Distributed Learning

Authors:[Hao Wu](https://arxiv.org/search/cs?searchtype=author&query=Wu,+H), [Kin Whye Chew](https://arxiv.org/search/cs?searchtype=author&query=Chew,+K+W), [Yizhan Han](https://arxiv.org/search/cs?searchtype=author&query=Han,+Y), [Han Li](https://arxiv.org/search/cs?searchtype=author&query=Li,+H), [Jingxian Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+J)

[View PDF](https://arxiv.org/pdf/2608.24516) [HTML (experimental)](https://arxiv.org/html/2608.24516v1)

> Abstract:Satellite-based distributed learning promises to train machine-learning models directly in orbit using massive, globally dispersed sensor data, thereby avoiding large-scale data downloads to ground servers. However, training convergence is significantly slowed by severe non-IID data, specifically label imbalance, as each satellite observes different geographic regions with distinct labels. This imbalance extends training duration and increases energy consumption for solar-powered satellites. Existing approaches either fully redistribute data to enforce IID conditions - accelerating convergence but incurring substantial communication delays - or avoid redistribution entirely by modifying local learning algorithms to mitigate the impact of label imbalance, which, however, still prolong training and increase energy use. Both extremes result in excessive total end-to-end learning time (data-transfer delay plus training time) and thus elevated onboard energy consumption.  
> We present SatDL, a data-redistribution framework designed to minimize total end-to-end learning time. At its core, SatDL develops a Distributor-Critic framework that jointly models and optimizes data-transfer delay and training time. Evaluations through trace-driven simulations of a 1,584-satellite Starlink constellation and hardware emulations using NVIDIA Jetson and A100 GPUs across five datasets show SatDL reduces total end-to-end learning time by up to 18.6% and onboard energy consumption by 12.23-88.00%, while maintaining inference accuracy within a few percentage points of state-of-the-art baselines.

| Comments: |  |
| --- | --- |
| Subjects: | Distributed, Parallel, and Cluster Computing (cs.DC); Machine Learning (cs.LG) |
| Cite as: | [arXiv:2608.24516](https://arxiv.org/abs/2608.24516) \[cs.DC\] |
|  | (or [arXiv:2608.24516v1](https://arxiv.org/abs/2608.24516v1) \[cs.DC\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.24516](https://doi.org/10.48550/arXiv.2608.24516) |

## Submission history

From: Kin Whye Chew \[[view email](https://arxiv.org/show-email/31a0169d/2608.24516)\]  
**\[v1\]** Tue, 25 Aug 2026 13:06:24 UTC (2,767 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.24516) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
