Search papers, labs, and topics across Lattice.
This paper introduces WAT3R, a novel feedforward framework designed to tackle the challenges of underwater 3D reconstruction caused by light attenuation and backscattering. By employing a lightweight neural adaptation module that adapts to underwater imaging degradation, WAT3R enhances multi-view reconstruction quality and outputs pixel-aligned 3D point maps and camera poses in a single forward pass. Experimental results demonstrate that WAT3R significantly outperforms existing state-of-the-art methods across multiple datasets, including FLSea, SQUID, and USOD10K, in tasks such as multi-view and monocular depth estimation.
WAT3R achieves high-quality underwater 3D reconstruction by effectively adapting to imaging degradation in a single forward pass, outperforming existing methods.
Reliable feedforward underwater 3D reconstruction remains challenging due to severe light attenuation and backscattering, which degrade visual quality and disrupt feature consistency across views, leading to inaccurate multi-view geometry. To address this issue, we propose WAT3R, a feed-forward framework for reconstructing 3D scenes directly from underwater images. By leveraging degradation adaptation as a geometry-constrained process, WAT3R integrates a lightweight neural adaptation module to flexibly account for these underwater imaging effects, thereby improving multi-view reconstruction quality. Implemented in a single forward pass, WAT3R directly and efficiently outputs pixel-aligned 3D point maps and camera poses from underwater videos, allowing a high-quality underwater 3D reconstruction. Experiments conducted on the FLSea, SQUID, and USOD10K datasets show that our method consistently outperforms state-of-the-art approaches on 3D reconstruction tasks, including multi-view/monocular depth estimation and camera pose estimation.