基于流表示实现人类演示的少样本模仿学习泛化 (Flow-Enabled Generalization to Human Demonstrations in Few-Shot Imitation Learning) - 专知论文

会员服务 ·

0

光流 · 演示 · 泛化 · 表示 · 视频 ·

Flow-Enabled Generalization to Human Demonstrations in Few-Shot Imitation Learning

翻译：基于流表示实现人类演示的少样本模仿学习泛化

Runze Tang,Penny Sweetser

from arxiv, Accepted to ICRA 2026

Imitation Learning (IL) enables robots to learn complex skills from demonstrations without explicit task modeling, but it typically requires large amounts of demonstrations, creating significant collection costs. Prior work has investigated using flow as an intermediate representation to enable the use of human videos as a substitute, thereby reducing the amount of required robot demonstrations. However, most prior work has focused on the flow, either on the object or on specific points of the robot/hand, which cannot describe the motion of interaction. Meanwhile, relying on flow to achieve generalization to scenarios observed only in human videos remains limited, as flow alone cannot capture precise motion details. Furthermore, conditioning on scene observation to produce precise actions may cause the flow-conditioned policy to overfit to training tasks and weaken the generalization indicated by the flow. To address these gaps, we propose SFCrP, which includes a Scene Flow prediction model for Cross-embodiment learning (SFCr) and a Flow and Cropped point cloud conditioned Policy (FCrP). SFCr learns from both robot and human videos and predicts any point trajectories. FCrP follows the general flow motion and adjusts the action based on observations for precision tasks. Our method outperforms SOTA baselines across various real-world task settings, while also exhibiting strong spatial and instance generalization to scenarios seen only in human videos.

翻译：模仿学习（Imitation Learning, IL）使机器人能够在不进行显式任务建模的情况下从演示中学习复杂技能，但其通常需要大量演示数据，导致高昂的采集成本。先前研究探索使用光流作为中间表示，以人类视频作为替代数据源，从而减少所需机器人演示的数量。然而，现有工作大多聚焦于物体或机器人/手部特定点上的光流，此类表示无法完整描述交互运动。同时，仅依赖光流实现对人类视频中观测场景的泛化能力仍有限制，因为单纯的光流无法捕捉精确的运动细节。此外，依赖场景观测生成精确动作可能导致基于光流的策略对训练任务过拟合，削弱光流所指示的泛化能力。为弥补这些不足，我们提出SFCrP方法，包含用于跨具身学习的场景流预测模型（SFCr）以及基于流与裁剪点云的条件策略（FCrP）。SFCr从机器人及人类视频中学习，并预测任意点的运动轨迹。FCrP遵循通用的流运动模式，并根据观测调整动作以执行精确任务。我们的方法在多种真实世界任务设定中均优于当前最先进的基线模型，同时对仅见于人类视频的场景展现出强大的空间与实例泛化能力。

0

相关内容

深度学习时代的模仿学习：新型分类体系与最新研究进展

深度学习时代的模仿学习：新型分类体系与最新研究进展

专知会员服务

11+阅读 · 2025年11月6日

机器人中的深度生成模型：多模态演示学习的综述

机器人中的深度生成模型：多模态演示学习的综述

专知会员服务

39+阅读 · 2024年8月9日

小样本学习报告：类人智能算法的初级形态，加速垂直场景下的AI普惠化

小样本学习报告：类人智能算法的初级形态，加速垂直场景下的AI普惠化

专知会员服务

42+阅读 · 2023年1月19日

模仿学习: 进展，分类和机会

专知会员服务

48+阅读 · 2021年7月2日

最新《模仿学习(Imitation Learning》进展报告, 加州理工Yisong Yue教授，附下载

最新《模仿学习(Imitation Learning》进展报告, 加州理工Yisong Yue教授，附下载

专知会员服务

41+阅读 · 2020年12月6日

【NeurIPS 2020】生成对抗性模仿学习的f-Divergence

【NeurIPS 2020】生成对抗性模仿学习的f-Divergence

专知会员服务

26+阅读 · 2020年10月9日

【复旦大学刘鹏飞博士论文】自然语言处理中的神经表示学习，153页pdf

专知会员服务

110+阅读 · 2020年9月1日

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning

专知会员服务

159+阅读 · 2020年2月29日

【CoRL2019最佳论文】模仿学习，A Divergence Minimization Perspective on Imitation Learning Methods

【CoRL2019最佳论文】模仿学习，A Divergence Minimization Perspective on Imitation Learning Methods

专知会员服务

24+阅读 · 2019年11月11日

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

专知会员服务

153+阅读 · 2019年10月18日

CVPR2020最新《小样本学习》综述教程，145页ppt带你学习最新FSL进展

CVPR2020最新《小样本学习》综述教程，145页ppt带你学习最新FSL进展

专知

40+阅读 · 2020年6月20日

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning，33页ppt

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning，33页ppt

专知

72+阅读 · 2020年2月29日

【Uber AI新论文】持续元学习，Learning to Continually Learn

【Uber AI新论文】持续元学习，Learning to Continually Learn

专知

19+阅读 · 2020年2月27日

【加州理工】什么是模仿学习(Imitation Learning（模仿学习), 这62页ppt带你了解进展，附下载

【加州理工】什么是模仿学习(Imitation Learning（模仿学习), 这62页ppt带你了解进展，附下载

专知

21+阅读 · 2019年11月14日

FewRel 2.0数据集：以近知远，以一知万，少次学习新挑战

FewRel 2.0数据集：以近知远，以一知万，少次学习新挑战

PaperWeekly

24+阅读 · 2019年11月6日

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

专知

363+阅读 · 2019年4月12日

基于逆强化学习的示教学习方法综述

基于逆强化学习的示教学习方法综述

计算机研究与发展

16+阅读 · 2019年2月25日

这可能是「多模态机器学习」最通俗易懂的介绍

这可能是「多模态机器学习」最通俗易懂的介绍

计算机视觉life

113+阅读 · 2018年12月20日

网络表示学习领域（NRL/NE）必读论文汇总

网络表示学习领域（NRL/NE）必读论文汇总

AI科技评论

16+阅读 · 2018年2月18日

迁移学习在深度学习中的应用

迁移学习在深度学习中的应用

专知

24+阅读 · 2017年12月24日

基于人机交互的数据驱动式人群行为建模与仿真研究

国家自然科学基金

4+阅读 · 2015年12月31日

基于高斯过程模型的多示例多标记学习算法研究

国家自然科学基金

14+阅读 · 2015年12月31日

多标记文本数据流分类方法研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于多样化查询的多标记主动学习研究

国家自然科学基金

0+阅读 · 2015年12月31日

面向大数据的安全迁移学习方法

国家自然科学基金

31+阅读 · 2015年12月31日

面向异分布数据的主动学习方法

国家自然科学基金

12+阅读 · 2015年12月31日

面向大规模数据流的集成学习模型与方法研究

国家自然科学基金

5+阅读 · 2014年12月31日

基于逆向强化学习和人工智能的移动机器人自主学习方法研究

国家自然科学基金

12+阅读 · 2013年12月31日

数据和模型混合驱动的虚拟人群行为仿真技术研究及其在军事中的应用

国家自然科学基金

10+阅读 · 2011年12月31日

强化学习关键技术及其在机器人行为学习中的应用

国家自然科学基金

23+阅读 · 2009年12月31日

Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos

Arxiv

0+阅读 · 2月13日

Scaling Single Human Demonstrations for Imitation Learning using Generative Foundational Models

Arxiv

0+阅读 · 2月13日

Self-Augmented Robot Trajectory: Efficient Imitation Learning via Safe Self-augmentation with Demonstrator-annotated Precision

Arxiv

0+阅读 · 2月11日

RFS: Reinforcement learning with Residual flow steering for dexterous manipulation

Arxiv

0+阅读 · 2月3日

PRISM: Performer RS-IMLE for Single-pass Multisensory Imitation Learning

Arxiv

0+阅读 · 2月2日

RFS: Reinforcement learning with Residual flow steering for dexterous manipulation

Arxiv

0+阅读 · 2月2日

Temporally Coherent Imitation Learning via Latent Action Flow Matching for Robotic Manipulation

Arxiv

0+阅读 · 1月30日

ConceptACT: Episode-Level Concepts for Sample-Efficient Robotic Imitation Learning

Arxiv

0+阅读 · 1月23日

Learning Diverse Skills for Behavior Models with Mixture of Experts

Arxiv

0+阅读 · 1月18日

Interactive and Hybrid Imitation Learning: Provably Beating Behavior Cloning

Arxiv

0+阅读 · 1月13日

VIP会员

文章信息

相关主题

相关VIP内容

深度学习时代的模仿学习：新型分类体系与最新研究进展

深度学习时代的模仿学习：新型分类体系与最新研究进展

专知会员服务

11+阅读 · 2025年11月6日

机器人中的深度生成模型：多模态演示学习的综述

机器人中的深度生成模型：多模态演示学习的综述

专知会员服务

39+阅读 · 2024年8月9日

小样本学习报告：类人智能算法的初级形态，加速垂直场景下的AI普惠化

小样本学习报告：类人智能算法的初级形态，加速垂直场景下的AI普惠化

专知会员服务

42+阅读 · 2023年1月19日

模仿学习: 进展，分类和机会

专知会员服务

48+阅读 · 2021年7月2日

最新《模仿学习(Imitation Learning》进展报告, 加州理工Yisong Yue教授，附下载

最新《模仿学习(Imitation Learning》进展报告, 加州理工Yisong Yue教授，附下载

专知会员服务

41+阅读 · 2020年12月6日

【NeurIPS 2020】生成对抗性模仿学习的f-Divergence

【NeurIPS 2020】生成对抗性模仿学习的f-Divergence

专知会员服务

26+阅读 · 2020年10月9日

【复旦大学刘鹏飞博士论文】自然语言处理中的神经表示学习，153页pdf

专知会员服务

110+阅读 · 2020年9月1日

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning

专知会员服务

159+阅读 · 2020年2月29日

【CoRL2019最佳论文】模仿学习，A Divergence Minimization Perspective on Imitation Learning Methods

【CoRL2019最佳论文】模仿学习，A Divergence Minimization Perspective on Imitation Learning Methods

专知会员服务

24+阅读 · 2019年11月11日

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

专知会员服务

153+阅读 · 2019年10月18日

热门VIP内容

开通专知VIP会员享更多权益服务

智能体记忆深度剖析：评价指标与系统局限性的分类体系及实证分析

《可信人工智能赋能系统的支柱》

【CMU博士论文】可靠轨迹预测的分层基石：数据、评估与方法

人工智能赋能边缘与自主系统：美陆军现代化进程聚焦威胁探测与战术边缘情报

相关资讯

CVPR2020最新《小样本学习》综述教程，145页ppt带你学习最新FSL进展

CVPR2020最新《小样本学习》综述教程，145页ppt带你学习最新FSL进展

专知

40+阅读 · 2020年6月20日

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning，33页ppt

元迁移学习的小样本学习，Meta-transfer Learning for Few-shot Learning，33页ppt

专知

72+阅读 · 2020年2月29日

【Uber AI新论文】持续元学习，Learning to Continually Learn

【Uber AI新论文】持续元学习，Learning to Continually Learn

专知

19+阅读 · 2020年2月27日

【加州理工】什么是模仿学习(Imitation Learning（模仿学习), 这62页ppt带你了解进展，附下载

【加州理工】什么是模仿学习(Imitation Learning（模仿学习), 这62页ppt带你了解进展，附下载

专知

21+阅读 · 2019年11月14日

FewRel 2.0数据集：以近知远，以一知万，少次学习新挑战

FewRel 2.0数据集：以近知远，以一知万，少次学习新挑战

PaperWeekly

24+阅读 · 2019年11月6日

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

《小样本学习(Few-shot learning)》最新41页综述论文，来自港科大和第四范式

专知

363+阅读 · 2019年4月12日

基于逆强化学习的示教学习方法综述

基于逆强化学习的示教学习方法综述

计算机研究与发展

16+阅读 · 2019年2月25日

这可能是「多模态机器学习」最通俗易懂的介绍

这可能是「多模态机器学习」最通俗易懂的介绍

计算机视觉life

113+阅读 · 2018年12月20日

网络表示学习领域（NRL/NE）必读论文汇总

网络表示学习领域（NRL/NE）必读论文汇总

AI科技评论

16+阅读 · 2018年2月18日

迁移学习在深度学习中的应用

迁移学习在深度学习中的应用

专知

24+阅读 · 2017年12月24日

相关论文

Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos

Arxiv

0+阅读 · 2月13日

Scaling Single Human Demonstrations for Imitation Learning using Generative Foundational Models

Arxiv

0+阅读 · 2月13日

Self-Augmented Robot Trajectory: Efficient Imitation Learning via Safe Self-augmentation with Demonstrator-annotated Precision

Arxiv

0+阅读 · 2月11日

RFS: Reinforcement learning with Residual flow steering for dexterous manipulation

Arxiv

0+阅读 · 2月3日

PRISM: Performer RS-IMLE for Single-pass Multisensory Imitation Learning

Arxiv

0+阅读 · 2月2日

RFS: Reinforcement learning with Residual flow steering for dexterous manipulation

Arxiv

0+阅读 · 2月2日

Temporally Coherent Imitation Learning via Latent Action Flow Matching for Robotic Manipulation

Arxiv

0+阅读 · 1月30日

ConceptACT: Episode-Level Concepts for Sample-Efficient Robotic Imitation Learning

Arxiv

0+阅读 · 1月23日

Learning Diverse Skills for Behavior Models with Mixture of Experts

Arxiv

0+阅读 · 1月18日

Interactive and Hybrid Imitation Learning: Provably Beating Behavior Cloning

Arxiv

0+阅读 · 1月13日

相关基金

基于人机交互的数据驱动式人群行为建模与仿真研究

国家自然科学基金

4+阅读 · 2015年12月31日

基于高斯过程模型的多示例多标记学习算法研究

国家自然科学基金

14+阅读 · 2015年12月31日

多标记文本数据流分类方法研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于多样化查询的多标记主动学习研究

国家自然科学基金

0+阅读 · 2015年12月31日

面向大数据的安全迁移学习方法

国家自然科学基金

31+阅读 · 2015年12月31日

面向异分布数据的主动学习方法

国家自然科学基金

12+阅读 · 2015年12月31日

面向大规模数据流的集成学习模型与方法研究

国家自然科学基金

5+阅读 · 2014年12月31日

基于逆向强化学习和人工智能的移动机器人自主学习方法研究

国家自然科学基金

12+阅读 · 2013年12月31日

数据和模型混合驱动的虚拟人群行为仿真技术研究及其在军事中的应用

国家自然科学基金

10+阅读 · 2011年12月31日

强化学习关键技术及其在机器人行为学习中的应用

国家自然科学基金

23+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员