Answering from Sure to Uncertain: Uncertainty-Aware Curriculum Learning for Video Question Answering

While significant advancements have been made in video question answering (VideoQA), the potential benefits of enhancing model generalization through tailored difficulty scheduling have been largely overlooked in existing research. This paper seeks to bridge that gap by incorporating VideoQA into a curriculum learning (CL) framework that progressively trains models from simpler to more complex data. Recognizing that conventional self-paced CL methods rely on training loss for difficulty measurement, which might not accurately reflect the intricacies of video-question pairs, we introduce the concept of uncertainty-aware CL. Here, uncertainty serves as the guiding principle for dynamically adjusting the difficulty. Furthermore, we address the challenge posed by uncertainty by presenting a probabilistic modeling approach for VideoQA. Specifically, we conceptualize VideoQA as a stochastic computation graph, where the hidden representations are treated as stochastic variables. This yields two distinct types of uncertainty: one related to the inherent uncertainty in the data and another pertaining to the model's confidence. In practice, we seamlessly integrate the VideoQA model into our framework and conduct comprehensive experiments. The findings affirm that our approach not only achieves enhanced performance but also effectively quantifies uncertainty in the context of VideoQA.

翻译：尽管视频问答（VideoQA）领域已取得显著进展，但现有研究在很大程度上忽略了通过针对性难度调度增强模型泛化能力的潜在益处。本文旨在通过将VideoQA融入课程学习（CL）框架来弥合这一差距，该框架从简单到复杂的数据逐步训练模型。考虑到传统自步课程学习方法依赖训练损失进行难度度量，这可能无法准确反映视频-问题对的复杂性，我们引入了不确定性感知的课程学习概念。其中，不确定性作为动态调整难度的指导原则。此外，我们通过提出VideoQA的概率建模方法来解决不确定性带来的挑战。具体而言，我们将VideoQA概念化为一个随机计算图，其中隐藏表示被视为随机变量。这产生了两种不同类型的不确定性：一种与数据固有的不确定性相关，另一种与模型自身的置信度相关。在实践中，我们将VideoQA模型无缝集成到我们的框架中，并进行了全面实验。实验结果证实，我们的方法不仅实现了性能提升，还有效量化了VideoQA背景下的不确定性。

相关内容

MoDELS

关注 45

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

O’Reilly报告：知识图谱崛起——面向现代数据集成和数据结构体系，“The Rise of the Knowledge Graph——Toward Modern Data Integration and the Data Fabric Architecture”

专知会员服务

49+阅读 · 2022年2月18日

UCM《机器学习导论笔记》，80页pdf CSE176 Introduction to Machine Learning

专知会员服务

32+阅读 · 2021年9月29日

Linux导论，Introduction to Linux，96页ppt

专知会员服务

82+阅读 · 2020年7月26日

FlowQA: Grasping Flow in History for Conversational Machine Comprehension

专知会员服务

35+阅读 · 2019年10月18日