---
格式版本: 2
标题: "rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation"
原文链接: "https://arxiv.org/abs/2608.18952"
发布日期: "2026-08-19"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Wed, 19 Aug 2026 14:17:34 UTC (9,573 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-20T15:52:07+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-20T15:48:16+08:00"
入库时间: "2026-08-20T07:52:08.047Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=model%20fit&searchtype=all"
匹配关键词:
  - "model fit"
  - "AI"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 12
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "内容为arXiv推荐系统论文，讨论LLM推理与记忆机制，与超节点/AI Rack/机柜级AI基础设施完全无关，仅因搜索词命中model fit。"
AI质检模型: "ali-deepseek-v4-flash"
AI质检时间: "2026-08-20T15:57:18+08:00"
AI主题相关性: 2
AI来源权威性: 4
AI新颖性: 2
AI技术细节: 2
AI商业部署信号: 1
AI完整性: 1
AI摘要: "rEDMRec将教师大模型的推荐推理蒸馏为四类可编辑经验记忆，由记忆控制器维护，学生模型仅检索该记忆即可排序，无需每轮重新调用大模型。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-07T03:20:06.429Z"
采集批次: "2026年8月20日14点19分32秒"
采集批次ID: "20260820-141932-079"
去重键: "https://arxiv.org/abs/2608.18952"
---

## Computer Science > Information Retrieval

## Title:rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation

Authors:[Minh Hoang Nguyen](https://arxiv.org/search/cs?searchtype=author&query=Nguyen,+M+H), [Tung Le](https://arxiv.org/search/cs?searchtype=author&query=Le,+T), [Huy Tien Nguyen](https://arxiv.org/search/cs?searchtype=author&query=Nguyen,+H+T)

[View PDF](https://arxiv.org/pdf/2608.18952) [HTML (experimental)](https://arxiv.org/html/2608.18952v1)

> Abstract:Large language models can improve recommendation quality by reasoning explicitly over user history and candidate items - for example, extracting a user's preferences or explaining why one item fits better than another - rather than mapping history directly to a ranked list. This reasoning, however, is expensive to repeat on every ranking request and, once produced, is typically consumed once and discarded, leaving it neither reusable across future requests nor easy to inspect or correct as user tastes drift. Our insight is that reasoning does not need to be regenerated at every call if it can instead be compressed once into a compact, structured memory that a lightweight model retrieves from. We propose rEDMRec, which distills a teacher LLM's reasoning into four typed, editable experience channels - long-term preference, short-term context, item-perception, and counterfactual hard-negative comparisons - maintained by an LLM memory controller that performs Add/Delete/Modify/Keep operations and refines entries via K-agent debate. A lightweight student LLM then ranks candidates purely by retrieving from this memory, without invoking the teacher again, decoupling online inference cost from reasoning depth. Across ML-1M, Amazon Beauty, and Steam and ten student backbones, rEDMRec improves HR@1 over zero-shot, few-shot, and RAG on every backbone, and over GraphRAG on most backbones, with Impv up to 13.3% vs. the second-best baseline on ML-1M. Channel ablations show that short-term context is the only channel that helps consistently across capacity tiers, whereas long-term, item-perception, and counterfactual contributions are capacity-dependent (and can reverse on the strongest students); debate-based memory optimization lowers bank duplication by 7.4 percentage points while raising downstream HR@1 by up to +0.029 over six optimization epochs.

| Subjects: | Information Retrieval (cs.IR); Artificial Intelligence (cs.AI); Computation and Language (cs.CL) |
| --- | --- |
| MSC classes: | 68T50, 68P20 |
| ACM classes: | H.3.3; I.2.6; I.2.7 |
| Cite as: | [arXiv:2608.18952](https://arxiv.org/abs/2608.18952) \[cs.IR\] |
|  | (or [arXiv:2608.18952v1](https://arxiv.org/abs/2608.18952v1) \[cs.IR\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.18952](https://doi.org/10.48550/arXiv.2608.18952) |

## Submission history

From: Minh Nguyen \[[view email](https://arxiv.org/show-email/6aad17b7/2608.18952)\]  
**\[v1\]** Wed, 19 Aug 2026 14:17:34 UTC (9,573 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.18952) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
