Recent years have witnessed a rapid development of immersive multimedia which bridges the gap between the real world and virtual space. Volumetric videos, as an emerging representative 3D video paradigm that empowers extended reality, stand out to provide unprecedented immersive and interactive video watching experience. Despite the tremendous potential, the research towards 3D volumetric video is still in its infancy, relying on sufficient and complete datasets for further exploration. However, existing related volumetric video datasets mostly only include a single object, lacking details about the scene and the interaction between them. In this paper, we focus on the current most widely used data format, point cloud, and for the first time release a full-scene volumetric video dataset that includes multiple people and their daily activities interacting with the external environments. Comprehensive dataset description and analysis are conducted, with potential usage of this dataset. The dataset and additional tools can be accessed via the following website: https://cuhksz-inml.github.io/full_scene_volumetric_video_dataset/.
翻译:近年来,沉浸式多媒体技术快速发展,弥合了现实世界与虚拟空间之间的鸿沟。体视频作为一种新兴的、赋能扩展现实的代表性三维视频范式,提供了前所未有的沉浸式和交互式视频观看体验。尽管潜力巨大,但三维体视频的研究仍处于起步阶段,需要足够且完整的数据集进行进一步探索。然而,现有的相关体视频数据集大多仅包含单个物体,缺乏场景细节及物体间的交互信息。本文聚焦于当前最广泛使用的数据格式——点云,并首次发布了一个包含多人及其与外部环境进行日常活动交互的全场景体视频数据集。我们对该数据集进行了全面的描述与分析,并探讨了其潜在应用。该数据集及附加工具可通过以下网站获取:https://cuhksz-inml.github.io/full_scene_volumetric_video_dataset/。