Multi-Object Self-Supervised Depth Denoising

Depth cameras are frequently used in robotic manipulation, e.g. for visual servoing. The quality of small and compact depth cameras is though often not sufficient for depth reconstruction, which is required for precise tracking in and perception of the robot's working space. Based on the work of Shabanov et al. (2021), in this work, we present a self-supervised multi-object depth denoising pipeline, that uses depth maps of higher-quality sensors as close-to-ground-truth supervisory signals to denoise depth maps coming from a lower-quality sensor. We display a computationally efficient way to align sets of two frame pairs in space and retrieve a frame-based multi-object mask, in order to receive a clean labeled dataset to train a denoising neural network on. The implementation of our presented work can be found at https://github.com/alr-internship/self-supervised-depth-denoising.

Paper

References (10)

09Selfsupervised depth denoising using lower-and higher-quality rgb-d sensors,2020 · International Conference on 3D Vision (3DV). IEEE,
10Dbscan revisited, revisited: why and how you should (still) use dbscan,2017 · ACM Transactions on Database Systems (TODS),

Similar papers

© 2026 NYSGPT2525 LLC