---
格式版本: 2
标题: "UBASE: An AI Search Engine for Trillion-Scale Vector Data Management at ByteDance"
原文链接: "https://arxiv.org/abs/2608.30607"
发布日期: "2026-08-31"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Mon, 31 Aug 2026 11:19:16 UTC (785 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-09-02T03:50:16+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-09-02T03:49:17+08:00"
入库时间: "2026-09-01T19:50:16.213Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=Scale-up&searchtype=all"
匹配关键词:
  - "Scale-up"
  - "deployment"
  - "latency"
  - "throughput"
  - "AI"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 18
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "论文涉及向量数据库与AI检索，未讨论超节点/AI Rack/机柜级系统、供电散热互连或相关硬件，与项目主题无关。"
AI质检模型: "zj-deepseek-v4-flash"
AI质检时间: "2026-09-02T03:50:40+08:00"
AI主题相关性: 0
AI来源权威性: 10
AI新颖性: 2
AI技术细节: 2
AI商业部署信号: 0
AI完整性: 4
AI摘要: "字节跳动论文介绍其统一 AI 搜索系统 UBASE，已支撑万亿级向量检索，最大部署索引近万亿高维向量。该系统通过量化感知内核与混合存储引擎，使吞吐最高提升 3 倍，索引内存降低 80%。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-01T23:06:25.942Z"
采集批次: "2026年9月1日23点56分52秒"
采集批次ID: "20260901-235652-761"
去重键: "https://arxiv.org/abs/2608.30607"
---

## Computer Science > Databases

## Title:UBASE: An AI Search Engine for Trillion-Scale Vector Data Management at ByteDance

Authors:[Yao Tian](https://arxiv.org/search/cs?searchtype=author&query=Tian,+Y), [Yuncheng Lu](https://arxiv.org/search/cs?searchtype=author&query=Lu,+Y), [Liyao Xiong](https://arxiv.org/search/cs?searchtype=author&query=Xiong,+L), [Yuming Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+Y), [Hao Zhang](https://arxiv.org/search/cs?searchtype=author&query=Zhang,+H), [Weichen Zhao](https://arxiv.org/search/cs?searchtype=author&query=Zhao,+W), [Xi Zhao](https://arxiv.org/search/cs?searchtype=author&query=Zhao,+X), [Bo Kuang](https://arxiv.org/search/cs?searchtype=author&query=Kuang,+B), [Dongyu Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+D), [Jiehui Li](https://arxiv.org/search/cs?searchtype=author&query=Li,+J), [Yakun Li](https://arxiv.org/search/cs?searchtype=author&query=Li,+Y), [Lei Zhang](https://arxiv.org/search/cs?searchtype=author&query=Zhang,+L)

[View PDF](https://arxiv.org/pdf/2608.30607) [HTML (experimental)](https://arxiv.org/html/2608.30607v1)

> Abstract:Since 2016, UBASE has been the foundation of ByteDance's search infrastructure, scaling to more than 7,000 clusters and 300 PB of indexed data. Driven by the demands of AI workloads, UBASE has evolved from a text search engine into a unified AI search system supporting vector retrieval, lexical matching, and predicate filtering. Its largest deployment indexes nearly one trillion high-dimensional vectors. This scale exposes two central bottlenecks in AI-era retrieval: memory-intensive graph-index construction under sustained ingestion, and the prohibitive cost of keeping vector indexes entirely in memory. UBASE addresses these bottlenecks with two techniques. First, it introduces a quantization-aware vector kernel based on SymRaBitQ, a new symmetric quantization scheme with tight theoretical guarantees that allows index construction to run directly in the quantized space accurately and efficiently without retaining a copy of full-precision vectors. Second, it provides a hybrid storage engine that supports memory-resident, hybrid, and SSD-resident deployments, with fine-grained record-level caching to trade memory for latency under operational control. On large-scale benchmarks, UBASE improves throughput by up to 3x, reduces indexing memory by 80%, and lowers operating cost by 86% compared with prior systems, while supporting trillion-vector scale, write-heavy or latency-sensitive workloads in production.

| Subjects: | Databases (cs.DB) |
| --- | --- |
| Cite as: | [arXiv:2608.30607](https://arxiv.org/abs/2608.30607) \[cs.DB\] |
|  | (or [arXiv:2608.30607v1](https://arxiv.org/abs/2608.30607v1) \[cs.DB\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.30607](https://doi.org/10.48550/arXiv.2608.30607) |

## Submission history

From: Yao Tian \[[view email](https://arxiv.org/show-email/8101f5f1/2608.30607)\]  
**\[v1\]** Mon, 31 Aug 2026 11:19:16 UTC (785 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.30607) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
