---
格式版本: 2
标题: "Behaviorally Effective LoRA Writes Are Sparse and Structured"
原文链接: "https://arxiv.org/abs/2609.01374"
发布日期: "2026-09-01"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Tue, 1 Sep 2026 15:09:43 UTC (1,385 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-09-02T18:52:23+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-09-02T18:51:39+08:00"
入库时间: "2026-09-02T10:52:23.750Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=Scale-up&searchtype=all"
匹配关键词:
  - "Scale-up"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 33
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文主线是LoRA写入子空间的稀疏性与结构性研究，新增可核验事实包括14次参数化切换、最高0.25%相对Frobenius误差及多个数据集上的top-k实验，但均属于模型适配与行为分析，不涉及机架级AI系统、互连、供电、液冷、RAS或可复用基础设施机制。来源为arXiv预印本摘要页，尚非正式同行评审成果；固定知识库未显示同一研究，但未命中不能证明首次出现。无客户、量产或生产部署信号，命中“应用与模型效率”硬否决项。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-09-02T18:53:26+08:00"
AI主题相关性: 1
AI来源权威性: 9
AI新颖性: 10
AI技术细节: 6
AI商业部署信号: 0
AI完整性: 7
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.81"
AI评分知识库SHA256: "3b93d12e47749b3512f545f51c44c011bdc0931677c2cfe61e4df59a2b1a5a48"
AI评分知识库检索词: "[\"Scale-up\",\"PDF\",\"arxiv.org/pdf/2609.01374\",\"HTML\",\"arxiv.org/html/2609.01374v1\",\"PCA\",\"GSM8K\",\"AQuA\",\"top-16\",\"top-32\",\"GSM8K/Qwen\",\"arxiv.org/abs/2609.01374\"]"
AI评分知识库命中: "[{\"id\":\"runtime-569639751a0dbe7ef3ffccc5\",\"title\":\"[2608.17503] Predict Before Replay: Joint FEC and Flight Control for Reliable Scale-Up Links\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-18\",\"matchedTerms\":[\"Scale-up\",\"PDF\",\"HTML\"],\"rank\":-10.41584664390761},{\"id\":\"july-correct-0111\",\"title\":\"StrataCL: Fabric-Native Communication Library for Production Supernodes\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"PDF\",\"HTML\"],\"rank\":-6.721112210816699},{\"id\":\"historical-jun-010\",\"title\":\"爱建证券-电子行业专题报告：Vera Rubin量产提速，RTX Spark打开终端AI新空间-260608.pdf\",\"sourceType\":\"curated_item\",\"time\":\"2026-06\",\"matchedTerms\":[\"PDF\"],\"rank\":-5.954285493762395},{\"id\":\"runtime-7aadc05d024a3a525d014ae9\",\"title\":\"MTIA 300: Meta’s First Training Chip with Built-in NICs and Communication-Offloading Engines\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-24\",\"matchedTerms\":[\"Scale-up\",\"PDF\"],\"rank\":-5.201441317874702},{\"id\":\"july-correct-0131\",\"title\":\"DeepSeek-V4如何在昇腾超节点高效完成全参数后训练？SLAI T-Rex技术报告解读\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"PDF\",\"HTML\",\"PCA\"],\"rank\":-5.156482840530344}]"
AI摘要: "研究显示，行为有效的 LoRA 写入是稀疏且结构化的。作者提出 Learned-Basis LoRA 方法，在 14 次切换中保持准确率不变，重构误差不超过 0.25%，并在 GSM8K、MathQA。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-02T17:13:27.593Z"
采集批次: "2026年9月2日18点28分48秒"
采集批次ID: "20260902-182848-894"
去重键: "https://arxiv.org/abs/2609.01374"
---

## Computer Science > Computation and Language

## Title:Behaviorally Effective LoRA Writes Are Sparse and Structured

Authors:[Haruto Sato](https://arxiv.org/search/cs?searchtype=author&query=Sato,+H), [Yuki Tanaka](https://arxiv.org/search/cs?searchtype=author&query=Tanaka,+Y), [Ren Nakamura](https://arxiv.org/search/cs?searchtype=author&query=Nakamura,+R), [Aoi Kobayashi](https://arxiv.org/search/cs?searchtype=author&query=Kobayashi,+A), [Mei Ito](https://arxiv.org/search/cs?searchtype=author&query=Ito,+M)

[View PDF](https://arxiv.org/pdf/2609.01374) [HTML (experimental)](https://arxiv.org/html/2609.01374v1)

> Abstract:Low-rank adaptation fixes the rank of the update, but it does not identify which parts of a trained  
> write actually carry behavior. We study that question directly and show that behaviorally effective  
> LoRA writes are sparse, structured, and far more concentrated than the raw low-rank parameterization  
> suggests.  
> We use Learned-Basis LoRA, a learned-basis continuation recipe, to expose that structure. The recipe  
> warms up an unconstrained adapter, converts its learned write columns into a module-wise orthonormal  
> basis, freezes that basis, and continues training inside the constrained parameterization. Across 14  
> exact switches from unconstrained to constrained form, held-out accuracy is unchanged at the  
> conversion step and reconstructed write matrices differ by at most 0.25% relative Frobenius error.  
> Same-state continuation then shows that the same trained checkpoint develops differently under  
> different write subspaces, establishing write geometry as a causal state variable. A no-retraining  
> projection test shows that useful write signal stays inside the learned write space and largely  
> disappears from random or frozen-activation PCA controls.  
> The concentration pattern is strong at both local and global scales. Across GSM8K, MathQA, and AQuA,  
> per-module top-k continuation reaches its optimum at k in {2, 4} in all twelve seed-level cases we  
> test. A stricter global ranking test shows that learned top-16 and top-32 subsets outperform matched  
> random subsets, especially on GSM8K/Qwen and MathQA/Qwen. Single-direction ablations further reveal a  
> sparse set of late q\_proj, o\_proj, and down\_proj components with outsized behavioral impact.

| Subjects: | Computation and Language (cs.CL) |
| --- | --- |
| Cite as: | [arXiv:2609.01374](https://arxiv.org/abs/2609.01374) \[cs.CL\] |
|  | (or [arXiv:2609.01374v1](https://arxiv.org/abs/2609.01374v1) \[cs.CL\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2609.01374](https://doi.org/10.48550/arXiv.2609.01374) |

## Submission history

From: Haruto Sato \[[view email](https://arxiv.org/show-email/a908cc4b/2609.01374)\]  
**\[v1\]** Tue, 1 Sep 2026 15:09:43 UTC (1,385 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2609.01374) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
