---
格式版本: 2
标题: "When Two Tracers Disagree: An Investigation of Multimodal Fusion for Clinical PET/CT Segmentation"
原文链接: "https://arxiv.org/abs/2608.19063"
发布日期: "2026-08-19"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Wed, 19 Aug 2026 16:04:46 UTC (8,657 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-20T15:31:50+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-20T15:26:21+08:00"
入库时间: "2026-08-20T07:31:50.727Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=performance&searchtype=all"
匹配关键词:
  - "performance"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 20
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "内容为医学图像分割论文，与超节点/AI Rack/机柜级AI基础设施完全无关，仅搜索词命中'performance'，非项目相关技术或商业信息。"
AI质检模型: "ali-deepseek-v4-flash"
AI质检时间: "2026-08-20T15:34:31+08:00"
AI主题相关性: 0
AI来源权威性: 10
AI新颖性: 2
AI技术细节: 0
AI商业部署信号: 0
AI完整性: 8
AI摘要: "该研究评估了PSMA与FDG PET/CT多模态融合在自动分割前列腺癌全身病灶中的效果，比较了早期融合和基于交叉注意力的中间融合策略。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-07T03:18:58.699Z"
采集批次: "2026年8月20日14点19分32秒"
采集批次ID: "20260820-141932-079"
去重键: "https://arxiv.org/abs/2608.19063"
---

## Computer Science > Computer Vision and Pattern Recognition

## Title:When Two Tracers Disagree: An Investigation of Multimodal Fusion for Clinical PET/CT Segmentation

Authors:[Jack A. Johnson](https://arxiv.org/search/cs?searchtype=author&query=Johnson,+J+A), [Bartłomiej W. Papież](https://arxiv.org/search/cs?searchtype=author&query=Papie%C5%BC,+B+W)

[View PDF](https://arxiv.org/pdf/2608.19063) [HTML (experimental)](https://arxiv.org/html/2608.19063v1)

> Abstract:PSMA and FDG PET/CT visualise complementary biological information in prostate cancer. Combining both tracers could capture heterogeneous tumour phenotypes that may be missed by either alone, yet there is no consensus on effective deep learning architectures for fusing these modalities. We evaluated multimodal image-fusion strategies for automatic whole-body PET/CT lesion segmentation to estimate total tumour burden. Using the public DEEP-PSMA Challenge dataset, we trained tracer-specific 3D nnU-Net baselines and compared (i) early fusion with a single encoder and one decoder (OEOD) or two decoders (OETD), and (ii) intermediate fusion via a dual-encoder cross-attention U-Net (DECA-UNet). Tracer-specific baselines performed strongly (PSMA Dice = 0.93; FDG = 0.81). Fusion yielded mixed results: OEOD produced a combined Dice of 0.90 (on an easier, non-tracer-specific task), whilst the tracer-specific fusion models reached PSMA/FDG = 0.69/0.64 (OETD) and 0.76/0.57 (DECA-UNet). Whilst fusion often provided reasonable PSMA segmentation, FDG performance degraded and no strategy consistently exceeded the single-tracer baselines. Under the evaluated setting, tracer-specific models remain the stronger baseline; clinically useful gains from multimodal fusion will likely require architectures that better preserve tracer specific representations. Our code is available at: [this https URL](https://github.com/JackJ3636/DEEP_PSMA_code)

| Comments: |  |
| --- | --- |
| Subjects: | Computer Vision and Pattern Recognition (cs.CV) |
| Cite as: | [arXiv:2608.19063](https://arxiv.org/abs/2608.19063) \[cs.CV\] |
|  | (or [arXiv:2608.19063v1](https://arxiv.org/abs/2608.19063v1) \[cs.CV\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.19063](https://doi.org/10.48550/arXiv.2608.19063) |

## Submission history

From: Jack Johnson \[[view email](https://arxiv.org/show-email/51ad104f/2608.19063)\]  
**\[v1\]** Wed, 19 Aug 2026 16:04:46 UTC (8,657 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.19063) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
