---
格式版本: 2
标题: "Efficient Passive Acoustic Monitoring of Killer Whales Using a Two-Stage Detection and Ecotype Classification Cascade"
原文链接: "https://arxiv.org/abs/2609.01792"
发布日期: "2026-09-01"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Tue, 1 Sep 2026 19:03:10 UTC (2,472 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-09-04T01:41:19+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-09-04T01:40:33+08:00"
入库时间: "2026-09-03T17:41:19.990Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=NVIDIA&searchtype=all"
匹配关键词:
  - "deployment"
相关厂家:
  - "NVIDIA"
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI优质: "否"
AI打分: 10
AI分档: "非优质"
AI质检状态: "不通过"
AI打分理由: "论文主题为虎鲸声学监测，与超节点/AI Rack等AI基础设施完全无关，仅提及NVIDIA H100作为推理硬件，无任何相关技术或商业信息。"
AI质检模型: "zj-deepseek-v4-flash"
AI质检时间: "2026-09-04T01:41:34+08:00"
AI主题相关性: 0
AI来源权威性: 5
AI新颖性: 0
AI技术细节: 0
AI商业部署信号: 0
AI完整性: 5
AI摘要: "研究提出一种基于ResNet的两阶段级联模型，先检测虎鲸发声，再将高置信度样本分类为东太平洋五种生态型，用于濒危南方定居虎鲸的实时被动声学监测。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-04T00:17:23.735Z"
采集批次: "2026年9月3日22点43分34秒"
采集批次ID: "20260903-224334-406"
去重键: "https://arxiv.org/abs/2609.01792"
---

## Computer Science > Sound

## Title:Efficient Passive Acoustic Monitoring of Killer Whales Using a Two-Stage Detection and Ecotype Classification Cascade

[View PDF](https://arxiv.org/pdf/2609.01792) [HTML (experimental)](https://arxiv.org/html/2609.01792v1)

> Abstract:Passive acoustic monitoring of killer whales is particularly important for conservation of the endangered Southern Resident killer whale population, but requires accurate models that can operate in real time under severe class imbalance and deployment shift. We propose a lightweight ResNet-based two-stage cascade that first detects killer whale vocalizations and then classifies confident detections into five eastern North Pacific ecotypes, abstaining on ambiguous calls. We train and evaluate the pipeline on the DCLDE 2027 dataset, where the detector achieves 0.960 macro-F1 and the classifier 0.958, outperforming frozen Perch 2.0 embeddings on the five-ecotype benchmark. By separating detection from ecotype recognition, the end-to-end cascade improves seven-class macro-F1 from 0.919 for a single-stage model to 0.933, with the largest gain on the rare OKW ecotype. To assess transfer beyond the benchmark, we use active learning to adapt the Stage 1 to the acoustic environment of Puget Sound, WA, increasing killer whale detection F1 from 0.405 to 0.755 on manually verified detection windows. Finally, each stage processes a 3 s window in approximately 1.4 ms on an NVIDIA H100, enabling faster than real time inference. These results demonstrate that the proposed two-stage cascade pipeline enables reliable killer whale detection and classification, adaptation to new acoustic domains, and real-time monitoring for conservation applications.

| Subjects: | Sound (cs.SD); Computer Vision and Pattern Recognition (cs.CV) |
| --- | --- |
| Cite as: | [arXiv:2609.01792](https://arxiv.org/abs/2609.01792) \[cs.SD\] |
|  | (or [arXiv:2609.01792v1](https://arxiv.org/abs/2609.01792v1) \[cs.SD\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2609.01792](https://doi.org/10.48550/arXiv.2609.01792) |

## Submission history

From: Daniela Ruiz \[[view email](https://arxiv.org/show-email/5a903aa3/2609.01792)\]  
**\[v1\]** Tue, 1 Sep 2026 19:03:10 UTC (2,472 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2609.01792) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
