High-quality video datasets are foundational for training robust models in tasks like action recognition, phase detection, and event segmentation. However, many real-world video datasets suffer from annotation errors such as *mislabeling*, where segments are assigned incorrect class labels, and *disordering*, where the temporal sequence does not follow the correct progression. These errors are particularly harmful in phase-annotated tasks, where temporal consistency is critical. We propose a novel, model-agnostic method for detecting annotation errors by analyzing the Cumulative Sample Loss (CSL)--defined as the average loss a frame incurs when passing through model checkpoints saved across training epochs. This per-frame loss trajectory acts as a dynamic fingerprint of frame-level learnability. Mislabeled or disordered frames tend to show consistently high or irregular loss patterns, as they remain difficult for the model to learn throughout training, while correctly labeled frames typically converge to low loss early. To compute CSL, we train a video segmentation model and store its weights at each epoch. These checkpoints are then used to evaluate the loss of each frame in a test video. Frames with persistently high CSL are flagged as likely candidates for annotation errors, including mislabeling or temporal misalignment. Our method does not require ground truth on annotation errors and is generalizable across datasets. Experiments on EgoPER and Cholec80 demonstrate strong detection performance, effectively identifying subtle inconsistencies such as mislabeling and frame disordering. The proposed approach provides a powerful tool for dataset auditing and improving training reliability in video-based machine learning.


翻译:高质量视频数据集是训练动作识别、阶段检测和事件分割等任务中鲁棒模型的基础。然而,许多现实世界视频数据集存在标注错误,例如*误标*(片段被分配了错误的类别标签)和*时序错乱*(时间序列未遵循正确的进展顺序)。这些错误在阶段标注任务中尤为有害,因为此类任务对时序一致性要求极高。本文提出一种新颖的、与模型无关的标注错误检测方法,通过分析累积样本损失(CSL)——定义为视频帧在训练过程中各检查点通过时产生的平均损失——来实现检测。这种逐帧损失轨迹可作为帧级可学习性的动态指纹。误标或时序错乱的帧往往呈现持续高损失或不规则的损失模式,因为它们在训练全程都难以被模型学习;而正确标注的帧通常能早期收敛至低损失。为计算CSL,我们训练一个视频分割模型并在每个训练周期保存其权重。这些检查点随后被用于评估测试视频中每一帧的损失。具有持续高CSL的帧被标记为可能的标注错误候选,包括误标或时序错位。本方法无需标注错误的真实标签,且可跨数据集泛化。在EgoPER和Cholec80数据集上的实验展示了强大的检测性能,能有效识别误标和帧序列错乱等细微不一致性。所提出的方法为数据集审计和提升视频机器学习训练可靠性提供了有力工具。

0
下载
关闭预览

相关内容

基于深度学习的视频异常检测:综述
专知会员服务
28+阅读 · 2024年9月10日
基于深度学习的视频目标检测综述
专知会员服务
85+阅读 · 2021年5月19日
专知会员服务
101+阅读 · 2020年7月20日
CVPR 2019:精确目标检测的不确定边界框回归
AI科技评论
13+阅读 · 2019年9月16日
视频目标识别资源集合
专知
25+阅读 · 2019年6月15日
基于视频的目标检测的发展【附PPT与视频资料】
人工智能前沿讲习班
19+阅读 · 2018年12月14日
目标检测算法盘点(最全)
七月在线实验室
17+阅读 · 2018年4月27日
国家自然科学基金
3+阅读 · 2017年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
12+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
VIP会员
最新内容
驱动军事决策变革的顶尖人工智能指挥系统
专知会员服务
7+阅读 · 8月11日
非对称防御中的自组织临界性:俄乌战争
专知会员服务
10+阅读 · 8月10日
《战争中的大语言模型监管》
专知会员服务
13+阅读 · 8月10日
《边缘计算关键技术分析及美军作战实践应用》
边缘计算的军事应用
专知会员服务
12+阅读 · 8月9日
一种考虑资源机动性的武器目标分配混合算法
专知会员服务
13+阅读 · 8月8日
相关基金
国家自然科学基金
3+阅读 · 2017年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
12+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员