Neural Radiance Field (NeRF) has recently emerged as a powerful representation to synthesize photorealistic novel views. While showing impressive performance, it relies on the availability of dense input views with highly accurate camera poses, thus limiting its application in real-world scenarios. In this work, we introduce Sparse Pose Adjusting Radiance Field (SPARF), to address the challenge of novel-view synthesis given only few wide-baseline input images (as low as 3) with noisy camera poses. Our approach exploits multi-view geometry constraints in order to jointly learn the NeRF and refine the camera poses. By relying on pixel matches extracted between the input views, our multi-view correspondence objective enforces the optimized scene and camera poses to converge to a global and geometrically accurate solution. Our depth consistency loss further encourages the reconstructed scene to be consistent from any viewpoint. Our approach sets a new state of the art in the sparse-view regime on multiple challenging datasets.
翻译:神经辐射场(NeRF)近期作为一种能够合成逼真新视角的强大表征方法而兴起。尽管其展现出令人印象深刻的表现,但该方法依赖于密集输入视图和高精度相机姿态的可用性,这限制了其在真实场景中的应用。本文提出稀疏姿态调整辐射场(SPARF),以解决在仅需少量宽基线输入图像(低至3张)且相机姿态含噪声的情况下进行新视角合成的挑战。我们的方法利用多视图几何约束来联合学习NeRF并优化相机姿态。通过利用输入视图间提取的像素匹配,我们的多视图对应目标强制优化后的场景和相机姿态收敛到全局几何精确的解。此外,深度一致性损失进一步促使重建场景从任意视角保持一致性。在多个具有挑战性的数据集上,我们的方法在稀疏视图场景下达到了新的最优性能。