Creating sound for storytelling is crucial to establishing the environment in productions such as films, TV series and video games. This process often involves repeating, layering and recording real objects or using sound libraries, which can be time-consuming and repetitive. To address these challenges, procedural audio, also known as digital foley, offers a solution by allowing sound designers to quickly generate samples. Despite its efficiency, questions remain about the believability of synthetic samples compared to real ones. In our study, we compared synthetic samples generated by an online procedural engine and integrated them with both animated and live-action visuals. Our results indicate that procedural audio is highly effective and perceived as believable in drama and sci-fi scenes, particularly for sound models such as lasers, hits, air and rockets, whereas synthetic sounds weren't as believable in cartoon productions when representing everyday actions. Finally, we identified specific models that needed optimisation and highlighted audio features that needed improvement with feedback from audio professionals.
翻译:声音创作对于电影、电视剧和电子游戏等作品中的环境营造至关重要。这一过程通常涉及对真实物体的重复录制、分层叠加,或使用音效库,既耗时又重复。为应对这些挑战,程序化音频(亦称数字拟音)通过允许声音设计师快速生成样本提供了解决方案。尽管效率显著,但合成样本相较于真实样本的可信度仍存疑问。本研究将在线程序化引擎生成的合成样本与动画及实拍画面进行整合比较。结果表明,程序化音频在剧情片和科幻场景中具有极高效率且被感知为可信,尤以激光、打击、空气与火箭等声音模型表现突出;然而在卡通作品中再现日常动作时,合成声音的可信度有所不足。最后,我们识别出若干需要优化的特定模型,并依据音频专业人士的反馈指出了有待改进的音频特征。