SketchDynamics：探索自由手绘草图在动画生成中的动态意图表达 (SketchDynamics: Exploring Free-Form Sketches for Dynamic Intent Expression in Animation Generation) - 专知论文

会员服务 ·

0

交互 · 视频 · 绘制 · 内容创作 · 塑造 ·

SketchDynamics: Exploring Free-Form Sketches for Dynamic Intent Expression in Animation Generation

翻译：SketchDynamics：探索自由手绘草图在动画生成中的动态意图表达

Boyu Li,Lin-Ping Yuan,Zeyu Wang,Hongbo Fu

from arxiv, conditionally accepted by CHI'26

Sketching provides an intuitive way to convey dynamic intent in animation authoring (i.e., how elements change over time and space), making it a natural medium for automatic content creation. Yet existing approaches often constrain sketches to fixed command tokens or predefined visual forms, overlooking their freeform nature and the central role of humans in shaping intention. To address this, we introduce an interaction paradigm where users convey dynamic intent to a vision-language model via free-form sketching, instantiated here in a sketch storyboard to motion graphics workflow. We implement an interface and improve it through a three-stage study with 24 participants. The study shows how sketches convey motion with minimal input, how their inherent ambiguity requires users to be involved for clarification, and how sketches can visually guide video refinement. Our findings reveal the potential of sketch and AI interaction to bridge the gap between intention and outcome, and demonstrate its applicability to 3D animation and video generation.

翻译：草图绘制为动画创作中的动态意图表达（即元素如何随时间与空间变化）提供了直观方式，使其成为自动化内容创作的自然媒介。然而，现有方法常将草图限制于固定指令标记或预定义视觉形式，忽略了其自由形式的本质以及人类在意图塑造中的核心作用。为此，我们提出一种交互范式：用户通过自由手绘草图向视觉-语言模型传达动态意图，并在此以草图故事板至动态图形的工作流程实现该范式。我们开发了交互界面，并通过包含24名参与者的三阶段研究对其改进。研究表明：草图如何以极简输入传达运动信息，其固有模糊性如何需要用户参与澄清，以及草图如何通过视觉引导视频优化。我们的发现揭示了草图与人工智能交互在弥合意图与结果之间鸿沟的潜力，并论证了其在三维动画与视频生成领域的适用性。

0

相关内容

【Nature Machine Intelligence】面向复杂系统建模的多模态图学习

【Nature Machine Intelligence】面向复杂系统建模的多模态图学习

专知会员服务

47+阅读 · 2023年12月19日

《视觉Transformer》最新简明综述，概述视觉Transformers 的不同架构设计和训练技巧

《视觉Transformer》最新简明综述，概述视觉Transformers 的不同架构设计和训练技巧

专知会员服务

67+阅读 · 2022年7月8日

【中科院自动化所】深度图生成方法及应用综述，A Survey on Deep Graph Generation: Methods and Applications

【中科院自动化所】深度图生成方法及应用综述，A Survey on Deep Graph Generation: Methods and Applications

专知会员服务

24+阅读 · 2022年3月15日

动态手势理解与交互综述

专知会员服务

34+阅读 · 2021年10月11日

图表示学习进展到哪了？看这份KDD2021《图表示学习:基础，方法，应用与系统》教程，众大牛讲解，附Slides

专知会员服务

67+阅读 · 2021年8月17日

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

专知会员服务

71+阅读 · 2021年3月7日

斯坦福大学李飞飞组发布Action Genome:一种新的表达形式，新的数据集，以及将动作分解成时空场景图的新模型

斯坦福大学李飞飞组发布Action Genome:一种新的表达形式，新的数据集，以及将动作分解成时空场景图的新模型

专知会员服务

40+阅读 · 2020年1月12日

【图机器学习论文】综述：图嵌入技术、应用和性能（Graph Embedding Techniques, Applications, and Performance: A Survey）

【图机器学习论文】综述：图嵌入技术、应用和性能（Graph Embedding Techniques, Applications, and Performance: A Survey）

专知会员服务

73+阅读 · 2019年12月16日

【WSDM 2020 论文】基于自关注网络的动态图表示学习（Dynamic graph representation learning via self-attention networks），Visa Research的研究员武延宏等

【WSDM 2020 论文】基于自关注网络的动态图表示学习（Dynamic graph representation learning via self-attention networks），Visa Research的研究员武延宏等

专知会员服务

98+阅读 · 2019年11月20日

【ACL 2019 Tutorials】基于图的含义表示:设计和处理（Graph-Based Meaning Representations: Design and Processing），Alexander Koller，Stephan Oepen，孙薇薇

【ACL 2019 Tutorials】基于图的含义表示:设计和处理（Graph-Based Meaning Representations: Design and Processing），Alexander Koller，Stephan Oepen，孙薇薇

专知会员服务

10+阅读 · 2019年11月16日

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

专知

27+阅读 · 2021年3月7日

图表示学习Graph Embedding综述

图表示学习Graph Embedding综述

图与推荐

10+阅读 · 2020年3月23日

【论文笔记】通过自注意力网络的动态图表示学习

【论文笔记】通过自注意力网络的动态图表示学习

专知

90+阅读 · 2019年12月2日

NLP+CV《桥接视觉与语言的研究综述》，带你全面了解视觉+语言最新应用和方法

NLP+CV《桥接视觉与语言的研究综述》，带你全面了解视觉+语言最新应用和方法

中国人工智能学会

27+阅读 · 2019年7月24日

图嵌入（Graph embedding）综述

图嵌入（Graph embedding）综述

人工智能前沿讲习班

449+阅读 · 2019年4月30日

图卷积网络介绍及进展【附PPT与视频资料】

图卷积网络介绍及进展【附PPT与视频资料】

人工智能前沿讲习班

24+阅读 · 2019年1月3日

图神经网络最近这么火，不妨看看我们精选的这七篇

图神经网络最近这么火，不妨看看我们精选的这七篇

人工智能前沿讲习班

37+阅读 · 2018年12月10日

自然语言处理中的自注意力机制（Self-Attention Mechanism）

自然语言处理中的自注意力机制（Self-Attention Mechanism）

PaperWeekly

22+阅读 · 2018年3月28日

【论文推荐】最新6篇图像描述生成相关论文—语言为枢纽、细粒度、生成器、注意力机制、策略梯度优化、判别性目标

【论文推荐】最新6篇图像描述生成相关论文—语言为枢纽、细粒度、生成器、注意力机制、策略梯度优化、判别性目标

专知

11+阅读 · 2018年3月20日

【论文推荐】最新5篇图像描述生成（Image Caption）相关论文—情感、注意力机制、遥感图像、序列到序列、深度神经结构

【论文推荐】最新5篇图像描述生成（Image Caption）相关论文—情感、注意力机制、遥感图像、序列到序列、深度神经结构

专知

66+阅读 · 2018年1月31日

基于单目RGB/RGBD相机的身体运动和面部运动同步捕获方法研究

国家自然科学基金

0+阅读 · 2017年12月31日

面向大类别的空中手写中英文识别技术研究

国家自然科学基金

2+阅读 · 2017年12月31日

“自然语言-草图”耦合的地理场景查询方法研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于内容感知编辑算子的复合型人脸图像真实感绘制

国家自然科学基金

0+阅读 · 2015年12月31日

仿动物大脑网格细胞神经定位机制的同步定位与地图构建方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

儿童手写运动促进中英文感知的认知神经机制

国家自然科学基金

0+阅读 · 2015年12月31日

基于草图语义部件的三维模型检索技术研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于草图的几何处理和应用

国家自然科学基金

2+阅读 · 2015年12月31日

中文句子语义概念图自动构建方法及应用研究

国家自然科学基金

3+阅读 · 2014年12月31日

烙画艺术模拟及其数字合成技术研究

国家自然科学基金

1+阅读 · 2014年12月31日

VideoSketcher: Video Models Prior Enable Versatile Sequential Sketch Generation

Arxiv

0+阅读 · 2月17日

Crane: An Accurate and Scalable Neural Sketch for Graph Stream Summarization

Arxiv

0+阅读 · 2月17日

SketchingReality: From Freehand Scene Sketches To Photorealistic Images

Arxiv

0+阅读 · 2月16日

Inspiration Seeds: Learning Non-Literal Visual Combinations for Generative Exploration

Arxiv

0+阅读 · 2月12日

Drawing Your Programs: Exploring the Applications of Visual-Prompting with GenAI for Teaching and Assessment

Arxiv

0+阅读 · 2月11日

Inspiration Seeds: Learning Non-Literal Visual Combinations for Generative Exploration

Arxiv

0+阅读 · 2月9日

Sketch2Scene: Automatic Generation of Interactive 3D Game Scenes from User's Casual Sketches

Arxiv

0+阅读 · 2月6日

IntentFlow: Investigating Fluid Dynamics of Intent Communication in Generative AI

Arxiv

0+阅读 · 1月28日

Automatic Synthesis of Visualization Design Knowledge Bases

Arxiv

0+阅读 · 1月27日

TechING: Towards Real World Technical Image Understanding via VLMs

Arxiv

0+阅读 · 1月26日

VIP会员

文章信息

相关主题

相关VIP内容

【Nature Machine Intelligence】面向复杂系统建模的多模态图学习

【Nature Machine Intelligence】面向复杂系统建模的多模态图学习

专知会员服务

47+阅读 · 2023年12月19日

《视觉Transformer》最新简明综述，概述视觉Transformers 的不同架构设计和训练技巧

《视觉Transformer》最新简明综述，概述视觉Transformers 的不同架构设计和训练技巧

专知会员服务

67+阅读 · 2022年7月8日

【中科院自动化所】深度图生成方法及应用综述，A Survey on Deep Graph Generation: Methods and Applications

【中科院自动化所】深度图生成方法及应用综述，A Survey on Deep Graph Generation: Methods and Applications

专知会员服务

24+阅读 · 2022年3月15日

动态手势理解与交互综述

专知会员服务

34+阅读 · 2021年10月11日

图表示学习进展到哪了？看这份KDD2021《图表示学习:基础，方法，应用与系统》教程，众大牛讲解，附Slides

专知会员服务

67+阅读 · 2021年8月17日

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

专知会员服务

71+阅读 · 2021年3月7日

斯坦福大学李飞飞组发布Action Genome:一种新的表达形式，新的数据集，以及将动作分解成时空场景图的新模型

斯坦福大学李飞飞组发布Action Genome:一种新的表达形式，新的数据集，以及将动作分解成时空场景图的新模型

专知会员服务

40+阅读 · 2020年1月12日

【图机器学习论文】综述：图嵌入技术、应用和性能（Graph Embedding Techniques, Applications, and Performance: A Survey）

【图机器学习论文】综述：图嵌入技术、应用和性能（Graph Embedding Techniques, Applications, and Performance: A Survey）

专知会员服务

73+阅读 · 2019年12月16日

【WSDM 2020 论文】基于自关注网络的动态图表示学习（Dynamic graph representation learning via self-attention networks），Visa Research的研究员武延宏等

【WSDM 2020 论文】基于自关注网络的动态图表示学习（Dynamic graph representation learning via self-attention networks），Visa Research的研究员武延宏等

专知会员服务

98+阅读 · 2019年11月20日

【ACL 2019 Tutorials】基于图的含义表示:设计和处理（Graph-Based Meaning Representations: Design and Processing），Alexander Koller，Stephan Oepen，孙薇薇

【ACL 2019 Tutorials】基于图的含义表示:设计和处理（Graph-Based Meaning Representations: Design and Processing），Alexander Koller，Stephan Oepen，孙薇薇

专知会员服务

10+阅读 · 2019年11月16日

热门VIP内容

开通专知VIP会员享更多权益服务

【CMU博士论文】基于自适应表征的高效视觉建模

《多域作战中融合网络、电子战与动能机动》

AI智能体时代大模型安全风险与攻防新挑战

迈向个性化大语言模型驱动的智能体：基础、评估与未来方向

相关资讯

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

【UCLA】动态图表示学习，40页ppt，Dynamic Graph Representation Learning

专知

27+阅读 · 2021年3月7日

图表示学习Graph Embedding综述

图表示学习Graph Embedding综述

图与推荐

10+阅读 · 2020年3月23日

【论文笔记】通过自注意力网络的动态图表示学习

【论文笔记】通过自注意力网络的动态图表示学习

专知

90+阅读 · 2019年12月2日

NLP+CV《桥接视觉与语言的研究综述》，带你全面了解视觉+语言最新应用和方法

NLP+CV《桥接视觉与语言的研究综述》，带你全面了解视觉+语言最新应用和方法

中国人工智能学会

27+阅读 · 2019年7月24日

图嵌入（Graph embedding）综述

图嵌入（Graph embedding）综述

人工智能前沿讲习班

449+阅读 · 2019年4月30日

图卷积网络介绍及进展【附PPT与视频资料】

图卷积网络介绍及进展【附PPT与视频资料】

人工智能前沿讲习班

24+阅读 · 2019年1月3日

图神经网络最近这么火，不妨看看我们精选的这七篇

图神经网络最近这么火，不妨看看我们精选的这七篇

人工智能前沿讲习班

37+阅读 · 2018年12月10日

自然语言处理中的自注意力机制（Self-Attention Mechanism）

自然语言处理中的自注意力机制（Self-Attention Mechanism）

PaperWeekly

22+阅读 · 2018年3月28日

【论文推荐】最新6篇图像描述生成相关论文—语言为枢纽、细粒度、生成器、注意力机制、策略梯度优化、判别性目标

【论文推荐】最新6篇图像描述生成相关论文—语言为枢纽、细粒度、生成器、注意力机制、策略梯度优化、判别性目标

专知

11+阅读 · 2018年3月20日

【论文推荐】最新5篇图像描述生成（Image Caption）相关论文—情感、注意力机制、遥感图像、序列到序列、深度神经结构

【论文推荐】最新5篇图像描述生成（Image Caption）相关论文—情感、注意力机制、遥感图像、序列到序列、深度神经结构

专知

66+阅读 · 2018年1月31日

相关论文

VideoSketcher: Video Models Prior Enable Versatile Sequential Sketch Generation

Arxiv

0+阅读 · 2月17日

Crane: An Accurate and Scalable Neural Sketch for Graph Stream Summarization

Arxiv

0+阅读 · 2月17日

SketchingReality: From Freehand Scene Sketches To Photorealistic Images

Arxiv

0+阅读 · 2月16日

Inspiration Seeds: Learning Non-Literal Visual Combinations for Generative Exploration

Arxiv

0+阅读 · 2月12日

Drawing Your Programs: Exploring the Applications of Visual-Prompting with GenAI for Teaching and Assessment

Arxiv

0+阅读 · 2月11日

Inspiration Seeds: Learning Non-Literal Visual Combinations for Generative Exploration

Arxiv

0+阅读 · 2月9日

Sketch2Scene: Automatic Generation of Interactive 3D Game Scenes from User's Casual Sketches

Arxiv

0+阅读 · 2月6日

IntentFlow: Investigating Fluid Dynamics of Intent Communication in Generative AI

Arxiv

0+阅读 · 1月28日

Automatic Synthesis of Visualization Design Knowledge Bases

Arxiv

0+阅读 · 1月27日

TechING: Towards Real World Technical Image Understanding via VLMs

Arxiv

0+阅读 · 1月26日

相关基金

基于单目RGB/RGBD相机的身体运动和面部运动同步捕获方法研究

国家自然科学基金

0+阅读 · 2017年12月31日

面向大类别的空中手写中英文识别技术研究

国家自然科学基金

2+阅读 · 2017年12月31日

“自然语言-草图”耦合的地理场景查询方法研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于内容感知编辑算子的复合型人脸图像真实感绘制

国家自然科学基金

0+阅读 · 2015年12月31日

仿动物大脑网格细胞神经定位机制的同步定位与地图构建方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

儿童手写运动促进中英文感知的认知神经机制

国家自然科学基金

0+阅读 · 2015年12月31日

基于草图语义部件的三维模型检索技术研究

国家自然科学基金

3+阅读 · 2015年12月31日

基于草图的几何处理和应用

国家自然科学基金

2+阅读 · 2015年12月31日

中文句子语义概念图自动构建方法及应用研究

国家自然科学基金

3+阅读 · 2014年12月31日

烙画艺术模拟及其数字合成技术研究

国家自然科学基金

1+阅读 · 2014年12月31日

微信扫码咨询专知VIP会员