---
格式版本: 2
标题: "ESA Unlocked: Tiered Cache - Break Through the Cache Hit Ratio Ceiling"
原文链接: "https://www.alibabacloud.com/blog/esa-unlocked-tiered-cache---break-through-the-cache-hit-ratio-ceiling_603514"
发布日期: "2026-08-31"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:configured_publication_date_rule"
发布时间证据: "alibaba-cloud-news-publication-date html:original: Bryan, Zhang August 31, 2026"
发布时间校准原因: "信源发布日期识别规则直接确认发布时间"
发布时间校准置信度: "high"
发布时间候选数量: 2
发布时间严格候选数量: 2
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-31T20:34:52+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-31T20:31:16+08:00"
入库时间: "2026-08-31T12:35:56.549Z"
来源平台: "固定入口"
搜索渠道: "fixed_url"
搜索词: "https://www.alibabacloud.com/blog"
匹配关键词:
  - "delivery"
  - "performance"
  - "latency"
  - "bandwidth"
相关厂家:
  - "阿里"
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
图片摘要:
  - "★ ./assets/img-f48e0e51.jpg | diagram | ESA多级缓存架构示意图，展示边缘、区域、智能缓存及Cache Reserve层级与回源路径。"
  - "✓ ./assets/img-ee607169.jpg | screenshot | ESA控制台Tiered Cache配置界面截图，展示功能入口与层级设置选项。"
  - "★ ./assets/img-72726609.jpg | chart | 单层缓存与多级缓存的命中率对比数据图，展示Tiered Cache带来的显著提升。"
  - "★ ./assets/img-a85bdc4d.jpg | diagram | ESA多级缓存工作流程图或配置示意图，体现各层级协同及请求处理逻辑。"
  - "★ ./assets/img-4036a8ef.jpg | diagram | Tiered Cache原理示意，对比传统单层与多级缓存的请求路径和缓存位置。"
AI摘要: "阿里云ESA推出多级缓存（Tiered Cache），在边缘节点外增加区域层、智能层和Cache Reserve持久层，可提升缓存命中率并减少回源，无需改代码或DNS。Cache Reserve为订阅制，每个实例支持最多20个网站；"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-08-31T23:37:05.552Z"
AI优质: "否"
AI打分: 33
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "正文主线是阿里云ESA多级CDN缓存的架构介绍、套餐选型与配置教程，并非超节点、AI Rack或机柜级AI基础设施。来源为厂商官方博客，内容完整，但未注明发布日期，也未明确披露新发布、GA、客户部署、实测指标或量产里程碑；所列场景明确说明并非实测。固定知识库未发现同产品历史记录，但不能据此认定首次出现。命中“教程与运维选型”硬否决项，且主题范围不匹配，当前页面不值得纳入超节点业务情报库。"
AI质检模型: "gpt-5.6-sol"
AI质检时间: "2026-09-01T07:40:20+08:00"
AI主题相关性: 0
AI来源权威性: 13
AI新颖性: 3
AI技术细节: 6
AI商业部署信号: 1
AI完整性: 10
AI评分提示词版本: "v17-精简生产版"
AI评分提示词SHA256: "48fb9777f386026761b4873eaff30807694fb11e9b352d7c69bf2dfde750cc7d"
AI评分知识库版本: "knowledge_base_v1-20260819+runtime.63"
AI评分知识库SHA256: "b2a4248d43868dee4c86d2b502d43668e260aea6b69d15f73a5dfe53287ca63f"
AI评分知识库检索词: "[\"阿里\",\"https://www.alibabacloud.com/blog\",\"RAS\",\"Intel\",\"ESA\",\"CDN\",\"NOT\",\"DNS\",\"OR\",\"HIT\",\"Xnip2026_08_31_12_01_15\",\"assets/img-f48e0e51.jpg\"]"
AI评分知识库命中: "[{\"id\":\"july-correct-0026\",\"title\":\"The Rackscale AI System Roadmaps That AMD Is Using To Chase Money\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"RAS\",\"Intel\",\"CDN\",\"NOT\",\"OR\",\"HIT\"],\"rank\":-8.88805619029553},{\"id\":\"runtime-0e8ec0dc686109af7d47b7ed\",\"title\":\"The Network Is More Of Nvidia’s Computer Than It Ever Was At Sun\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-27\",\"matchedTerms\":[\"RAS\",\"Intel\",\"CDN\",\"NOT\",\"OR\",\"HIT\"],\"rank\":-8.283770436356225},{\"id\":\"runtime-5581f72bc66fda8e6c6bf711\",\"title\":\"Advancing Standards-Based AI Fabric RAS Through COSMOS Integration\",\"sourceType\":\"ai_excellent_article\",\"time\":\"2026-08-10\",\"matchedTerms\":[\"RAS\",\"Intel\",\"NOT\",\"OR\",\"HIT\"],\"rank\":-8.201090867381557},{\"id\":\"july-correct-0081\",\"title\":\"AAI 2026: 6th Gen AMD EPYC Server CPUs Power the Agentic Data Center\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"RAS\",\"Intel\",\"NOT\",\"OR\",\"HIT\"],\"rank\":-8.097173409674358},{\"id\":\"july-correct-0062\",\"title\":\"Why AI Racks Need an Open Signal Conditioning Standard\",\"sourceType\":\"labeled_article\",\"time\":\"2026-07\",\"matchedTerms\":[\"RAS\",\"CDN\",\"NOT\",\"OR\",\"HIT\"],\"rank\":-7.240157304434924}]"
采集批次: "2026年8月31日20点21分03秒"
采集批次ID: "20260831-202103-550"
去重键: "https://www.alibabacloud.com/blog/esa-unlocked-tiered-cache---break-through-the-cache-hit-ratio-ceiling_603514"
---

Most traditional CDN users stop at single-tier edge caching and accept a modest cache hit ratio as "normal." However, It's NOT - it's a hard architectural ceiling. ESA's multi-level cache (Tiered Cache) adds Regional, Smart, and Cache Reserve layers that substantially lift hit ratios and offload your origin. No code changes. No DNS changes. Just enable the right tiers.

## What Is Tiered Cache?

Alibaba Cloud ESA (Edge Security Acceleration) provides a **multi-level caching architecture** that goes beyond traditional single-tier CDN caching. Instead of one shot at caching content at the edge, ESA gives every request **multiple chances** to be served from cache — across edge, regional, and origin-adjacent tiers, plus a persistent cold-storage layer — before it ever touches your origin server.

![Xnip2026_08_31_12_01_15](./assets/img-f48e0e51.jpg)

The result: a cache hit ratio well beyond the ceiling that single-tier edge caching imposes, and far fewer requests reaching your origin.

> **New to ESA?** ESA is Alibaba Cloud's edge security and acceleration platform, combining CDN, WAF, DDoS protection, and edge computing in one product. In the ESA console this feature lives under **Caching > Tiered Cache**.

![Xnip2026_08_31_11_52_03](./assets/img-ee607169.jpg)

This article explains the architecture, walks through where it helps most, and shows how to set it up.

---

## The Problem: Why Single-Tier Cache Has a Ceiling

Traditional CDN caching works in one layer: content gets cached at the edge POP closest to the user. If it's not there, the request goes all the way back to origin.

This works well for hot content: the small share of your assets that generates most of your traffic. But it fails for everything else:

| Problem | Impact |
| --- | --- |
| **Limited edge storage** | Edge POPs have limited capacity. Cold content gets evicted aggressively. |
| **No second chance** | Once evicted from an edge POP, the next request goes to origin — even if the same content is still cached in another region. |
| **Geographic fragmentation** | A file cached in one region doesn't help a user in another. Each edge POP operates independently. |
| **Cold content penalty** | Long-tail content (old blog posts, niche product images, archived docs) never stays cached long enough. |
| **The Pareto trap** | Most traffic concentrates on a small slice of content. That hot slice caches beautifully; the long tail generates nearly **all** cache misses. |

The result is predictable: most users enable CDN, watch their cache hit ratio climb to a modest level, and stop there. They scale origin servers and add bandwidth — working **around** the problem instead of solving it.

![Xnip2026_08_31_11_53_24](./assets/img-72726609.jpg)

**ESA's multi-level cache solves this.**

---

## How Tiered Cache Works: Pick the Right Topology

ESA's multi-level cache isn't a single fixed pipeline — it's a set of **tier configurations you choose between**, depending on your traffic pattern and plan. Each configuration adds an intermediate caching layer between the edge and your origin:

| Configuration | Tiers | What it does | Best for |
| --- | --- | --- | --- |
| **Edge Tiered Cache** | 1 | Edge nodes only (3,200+ globally), nearest to each user | Hot content, baseline caching |
| **Edge + Regional** | 2 | Regional nodes (deployed by region/province, near the edge) consolidate traffic from edge nodes in the same area before fetching from origin | Improving origin-fetch performance across a multi-region audience |
| **Edge + Smart** | 2 | Smart (intelligent) nodes sit **near your origin**, fewer of them, consolidating origin-fetch traffic | Maximizing hit ratio when you want to converge origin load |
| **All Cache Tiers** | 3 | Edge + Regional + Smart combined (Enterprise only) | High-traffic scenarios needing maximum origin convergence |

2 things to note:  
**Regional and Smart are alternatives, not sequential layers.** At the two-tier level you pick the topology that fits: Regional (nodes near the edge, organized by region) or Smart (nodes near the origin). You don't stack one on top of the other. Only the Enterprise **All Cache Tiers** mode combines edge, regional, and smart into a single three-tier path.

**Smart Tier is especially powerful for origins in mainland China.** If your origin is in Hangzhou, Smart Tier uses a nearby Shanghai-area POP as an intermediate cache — so an edge POP in Europe hits a Smart Tier POP in China instead of making a transcontinental request to your origin.

The full Enterprise path (All Cache Tiers + Cache Reserve) looks like this:

```
User Request
    │
    ▼
┌──────────────────────────────────────┐
│  Edge Tiered Cache                    │  3,200+ nodes · nearest to user
│  (Tokyo, Mumbai, Frankfurt, ...)      │  single-tier baseline
└──────────────┬───────────────────────┘
               │ miss
               ▼
┌──────────────────────────────────────┐
│  Regional Tiered Cache                │  near edge · by region/province
│  consolidates regional traffic        │  (Pro+ — OR Smart, pick one)
└──────────────┬───────────────────────┘
               │ miss
               ▼
┌──────────────────────────────────────┐
│  Smart Tiered Cache                   │  near origin · fewer nodes
│  converges origin fetch               │  (Pro+ — OR Regional, pick one)
└──────────────┬───────────────────────┘
               │ miss
               ▼
┌──────────────────────────────────────┐
│  Cache Reserve                        │  persistent cold-storage POP
│  (subscription instance)              │  no eviction · cold long-tail
└──────────────┬───────────────────────┘
               │ miss
               ▼
          Origin Server
```

> **Reading the diagram:** Pro and Premium plans choose **either** Regional **or** Smart as their second tier. The Enterprise **All Cache Tiers** mode is what wires edge + regional + smart into one three-tier path; Cache Reserve then sits on top as a persistent layer for cold content.

### Cache Reserve: A Persistent Layer for Cold Content

Cache Reserve is a separate, **persistent** caching layer for content that should be cached indefinitely — the long tail that regular tiers evict. ESA routes requests through a top-level Cache Reserve POP that holds less-popular files so they aren't pushed out by hot content.

Unlike edge/regional tiers that use LRU-style eviction, Cache Reserve keeps content until you purge it or your retention expires. It targets **cold traffic**; for hot traffic you typically add a cache rule to bypass Cache Reserve and save storage cost.

**Important:** Cache Reserve is **subscription-based**. You purchase a Cache Reserve instance (choosing a region and storage capacity) through your sales team, then associate websites with it. Each instance supports up to 20 websites, and it works best when the tiered cache mode is set to **All Cache Tiers** (other modes work but yield lower hit ratios and higher storage usage).

---

## Where It Helps Most: Representative Scenarios

> **Note:** These are representative scenarios illustrating where each tier adds value. They describe the pattern, not measured benchmarks.

![Xnip2026_08_31_11_54_54](./assets/img-a85bdc4d.jpg)

### Scenario 1: Gaming Platform with Frequent Patches

Game patches are large and short-lived at the edge — new patches evict old ones quickly, but plenty of users haven't updated yet. A **regional tier** holds patches far longer than any single edge POP, so users across the region who haven't updated still pull the patch from cache instead of the origin.

**The benefit:** far fewer origin fetches for large assets, faster patch downloads for laggards.

### Scenario 2: News Publisher with a Deep Archive

Breaking news is cached aggressively at the edge, but yesterday's articles get evicted as new stories arrive. A **regional tier** keeps recent articles warm much longer, and **Cache Reserve** holds the deep archive permanently.

**The benefit:** archive reads stay off the origin; you can run fewer origin servers.

### Scenario 3: High-Volume SaaS API

For an API serving many regions, most reads are cacheable but spread thin across edge POPs, so each POP misses on its own. **Smart Tier** converges origin-fetch traffic through a node near the origin, so repeated misses consolidate into far fewer origin requests.

**The benefit:** origin load converges; lower origin infrastructure and bandwidth cost.

### Scenario 4: E-commerce with a Massive SKU Catalog

The vast majority of product SKUs receive no views for long stretches — a classic long tail. Edge tiers evict them, so every rare view misses cache and hits the origin. **Cache Reserve** caches the entire catalog persistently.

**The benefit:** the whole long tail stays cached; origin load drops sharply.

### Scenario 5: Knowledge Base with Sporadic Spikes

Articles sit unviewed for weeks, then a social-media share triggers a traffic spike. Without a persistent layer, those articles always miss cache and hit the database. **Cache Reserve** keeps every article cached, so spikes are absorbed.

**The benefit:** database load stays flat even during viral spikes.

---

## How the Layers Compound

Each layer catches the requests the one before it missed:

- The **edge tier** serves the hot content.
- The **regional or smart tier** catches the warm content the edge evicted.
- **Cache Reserve** catches the cold long-tail the other tiers would eventually evict.

Stack them, and the share of requests that reach your origin shrinks at every step — often to a small fraction of total traffic. Your exact gain depends on your traffic shape (how much hot vs. cold content you serve), but the direction is consistent: **more tiers, fewer origin requests.**

The cost impact follows directly. Fewer origin requests means lower origin compute, storage, and bandwidth spend — and less need to over-provision origin capacity for traffic that a cache could have absorbed.

---

## Setup Guide

### Step 1: Enable Tiered Cache

ESA Console → **Websites** → click your site → **Caching > Tiered Cache** → **Configure**.

| Option | Plan | Best for |
| --- | --- | --- |
| Edge Tiered Cache | Entrance / Pro / Premium / Enterprise | Default single-layer caching |
| Edge + Regional | Pro / Premium / Enterprise | Multi-region audience, better origin-fetch performance |
| Edge + Smart | Pro / Premium / Enterprise | Origin in a specific geography (esp. mainland China), maximize hit ratio |
| All Cache Tiers | Enterprise | Maximum origin convergence (contact us to request) |

Pick one configuration and confirm. No DNS changes, no code changes, no origin reconfiguration. Settings take effect in about 3-5 minutes.

> **Heads-up:** When multi-level cache is set to **Edge Tiered Cache** only, the **Cache Warm-up** feature is unavailable. If you rely on warm-up, use a two-tier or higher configuration.

### Step 2: Set Up Cache Reserve (Enterprise)

Cache Reserve is a subscription instance, not a simple toggle:

1. **Purchase an instance** — contact your sales team to order a Cache Reserve instance for your desired region and storage capacity. ESA provisions it after purchase.
2. ESA Console → **Websites** → your site → **Caching > Cache Reserve** → **Configure**.
3. Turn on **Status**, then select your instance from the **Instance** list — choose the instance **closest to your origin** (e.g., origin in China (Hong Kong) → select the China (Hong Kong) instance).
4. Define what to cache (URL patterns like `/images/products/*`, `/docs/*`; extensions like `.jpg`, `.png`, `.pdf`; or use Cache Rules for complex logic) and set retention (indefinite or time-based).

Notes: each instance supports up to 20 websites; for best results set the tiered cache mode to **All Cache Tiers**; for hot traffic, add a cache rule to bypass Cache Reserve and reduce storage cost.

### Step 3: Validate with Response Headers

Check the `X-Cache` response header to confirm whether content was served from cache:

| Header | Meaning |
| --- | --- |
| `HIT` | Served from cache |
| `MISS` | Fetched from origin |

Give it a day or two, then check the cache hit ratio (with per-tier breakdown where available) in ESA analytics.

---

## Common Configuration Mistakes

Even with Tiered Cache enabled, these misconfigurations will reduce effectiveness:

| Mistake | Why it hurts | Fix |
| --- | --- | --- |
| **Short TTLs on static content** | A few minutes of `max-age` on product images means they expire constantly, defeating the regional/smart tier entirely | Set long TTLs (days/weeks) for static assets |
| **Overly aggressive cache purge** | Purging everything after a deploy wipes all tiers → thundering herd to origin | Purge selectively by URL pattern or tag |
| **Ignoring Vary header** | `Vary: User-Agent` creates a separate cache entry per user agent, fragmenting cache | Remove unnecessary Vary headers, or normalize with Transform Rules |
| **Cache key missing headers** | `/api/users` returns per-user data but the cache key omits the Authorization header | Include relevant headers in the cache key via Cache Rules |
| **Expecting warm-up on Edge-only mode** | Cache Warm-up is unavailable when multi-level cache is set to Edge Tiered Cache | Use a two-tier (Regional or Smart) or higher config if you need warm-up |
| **Sending hot traffic through Cache Reserve** | Cache Reserve is for cold content; caching hot content there wastes paid storage | Add a cache rule to bypass Cache Reserve for hot paths |
| **Skipping Cache Reserve for long-tail** | Regional/Smart tiers still evict eventually. Without Reserve, old product images still miss | Enable Cache Reserve for catalog/archive paths |

---

## Which Configuration Fits Your Use Case?

| Scenario | Recommended setup |
| --- | --- |
| Small site, mostly hot content | Edge Tiered Cache |
| E-commerce, moderate catalog | Edge + Regional |
| Origin in mainland China, global audience | Edge + Smart |
| Large media platform, long-tail content | All Cache Tiers + Cache Reserve |
| Software company, frequent releases | Edge + Smart + Cache Reserve |
| Gaming, massive asset library | All Cache Tiers + Cache Reserve |

![Xnip2026_08_31_11_58_44](./assets/img-4036a8ef.jpg)

**Rule of thumb:**

- **Hot content** (trending now) → Edge Tiered Cache is enough
- **Warm content** (popular last week/month), multi-region → add Regional
- **Converge origin load**, origin in a fixed geography → add Smart
- **Cold content** (long tail, archive) → add Cache Reserve
- **Maximum hit ratio** → All Cache Tiers + Cache Reserve (Enterprise)

---

## FAQ

**Does Tiered Cache add latency?**  
No, it usually reduces it. The edge POP still serves the user directly; the only difference is where the edge POP fetches content from on a miss. Smart Tier often lowers latency because its nodes sit closer to your origin than the edge POP would.

**Cache Reserve vs. OSS, what's the difference?**  
OSS is your origin storage — you upload and manage files there. Cache Reserve is a caching layer ESA manages automatically. Think of OSS as your warehouse; Cache Reserve as a shelf in the delivery truck that never gets cleared out.

**Do I need Cache Rules to use Tiered Cache?**  
No. Tiered Cache works automatically from your existing cache settings. Cache Rules just add per-path control — e.g. "cache this path at the regional tier for a long window, this other path at edge for a short one."

**Can I use Cache Reserve without Tiered Cache?**  
Cache Reserve works best with the tiered cache mode set to **All Cache Tiers**. Other modes work but yield lower hit ratios and higher storage usage.

**What plans include each tier?**

| Feature | Entrance | Pro | Premium | Enterprise |
| --- | --- | --- | --- | --- |
| Edge Tiered Cache | ✅ | ✅ | ✅ | ✅ |
| Edge + Regional | ❌ | ✅ | ✅ | ✅ |
| Edge + Smart | ❌ | ✅ | ✅ | ✅ |
| All Cache Tiers | ❌ | ❌ | ❌ | ✅ |
| Cache Reserve | ❌ | ❌ | ❌ | ✅ |

![Xnip2026_08_31_12_01_15](./assets/img-f48e0e51.jpg)

![Xnip2026_08_31_11_52_03](./assets/img-ee607169.jpg)

![Xnip2026_08_31_11_53_24](./assets/img-72726609.jpg)

![Xnip2026_08_31_11_54_54](./assets/img-a85bdc4d.jpg)

![Xnip2026_08_31_11_58_44](./assets/img-4036a8ef.jpg)
