SEABAD: A Tropical Bird Activity Detection Dataset for Passive Acoustic Monitoring

Passive acoustic monitoring (PAM) enables large-scale biodiversity assessment, but continuous recording generates large amounts of non-informative audio, creating challenges for storage, power consumption, and long-term edge deployment. Bird audio detection (BAD), which identifies bird vocalizations, can reduce this burden by filtering irrelevant recordings before downstream analysis. However, most BAD systems are trained on temperate datasets despite tropical soundscapes being denser, more species-rich, and acoustically unpredictable. To address this gap, we introduce SEABAD (Southeast Asian Bird Activity Detection), a dataset of 50,000 curated three-second clips from Southeast Asian soundscapes, evenly balanced between bird-present and bird-absent samples. The dataset spans 1,677 bird species and is standardized to 16 kHz mono audio for embedded and low-power inference. We developed a dual-branch curation pipeline: a six-stage positive-label workflow applied to Xeno-Canto recordings, alongside six source-specific negative-label extractions from environmental datasets. These procedures reduced class imbalance by 13.7% (Gini coefficient: 0.601 to 0.519). A manual audit of 1,000 positive clips confirmed 97.8% +/- 0.9% labeling accuracy. Baseline experiments using MobileNetV3-Small achieved 99.57% +/- 0.25% accuracy and 0.9985 +/- 0.0002 AUC across three random seeds. SEABAD and the full curation pipeline are publicly released to support tropical BAD research and energy-efficient acoustic monitoring.

翻译：被动声学监测（PAM）能够实现大规模生物多样性评估，但连续录音会产生大量非信息性音频，给存储、功耗及长期边缘部署带来挑战。鸟类音频检测（BAD）通过识别鸟鸣声，可在下游分析前过滤无关录音以减轻这一负担。然而，尽管热带声景密度更高、物种更丰富且声学不可预测性更强，现有BAD系统多基于温带数据集训练。为解决这一差距，我们提出SEABAD（东南亚鸟类活动检测）——一个包含50,000个精选三秒片段的数据集，样本源自东南亚声景，其中鸟类存在与不存在样本均衡分布。该数据集涵盖1,677种鸟类，并标准化为16kHz单声道音频以适配嵌入式及低功耗推理。我们开发了双分支数据筛选流程：一条六阶段正标签处理流水线针对Xeno-Canto录音，另一条六类特定来源负标签提取流水线从环境数据集中获取样本。这些流程将类别不平衡度降低13.7%（基尼系数从0.601降至0.519）。对1,000个正标签片段的人工审核确认标注准确率为97.8%±0.9%。基于MobileNetV3-Small的基线实验在三个随机种子上取得99.57%±0.25%的准确率与0.9985±0.0002的AUC值。SEABAD数据集及完整筛选流程已公开，以支持热带BAD研究与能效型声学监测。

相关内容

数据集

关注 88

数据集，又称为资料集、数据集合或资料集合，是一种由数据所组成的集合。
Data set（或dataset）是一个数据的集合，通常以表格形式出现。每一列代表一个特定变量。每一行都对应于某一成员的数据集的问题。它列出的价值观为每一个变量，如身高和体重的一个物体或价值的随机数。每个数值被称为数据资料。对应于行数，该数据集的数据可能包括一个或多个成员。

基于声学的无人机检测技术综述

专知会员服务

17+阅读 · 5月30日

《深度学习技术在海战舰船声景分类中的应用研究》最新63页

专知会员服务

28+阅读 · 2025年5月20日

《使用 "先跟踪后检测 "方法在主动声纳阵列中检测和跟踪多个移动目标》

专知会员服务

20+阅读 · 2024年4月24日

《利用音频传感器网络检测、识别和跟踪无人机的时频协同方法》

专知会员服务

41+阅读 · 2023年9月11日