Aberration-Aware Depth-from-Focus
Computer vision methods for depth estimation usually use simple camera models with idealized optics. For modern machine learning approaches, this creates an issue when attempting to train deep networks with simulated data, especially for focus-sensitive tasks like Depth-from-Focus. In this work, we investigate the domain gap caused by off-axis aberrations that will affect the decision of the best-focused frame in a focal stack. We then explore bridging this domain gap through aberration-aware training (AAT). Our approach involves a lightweight network that models lens aberrations at different positions and focus distances, which is then integrated into the conventional network training pipeline. We evaluate the generality of pretrained models on both synthetic and real-world data. Our experimental results demonstrate that the proposed AAT scheme can improve depth estimation accuracy without fine-tuning the model or modifying the network architecture.
Computer vision methods for depth estimation usually use simple camera models with idealized optics. For modern machine learning approaches, this creates an issue when attempting to train deep networks with simulated data, especially for focus-sensitive tasks like Depth-from-Focus. In this work, we investigate the domain gap caused by off-axis aberrations that will affect the decision of the best-focused frame in a focal stack. We then explore bridging this domain gap through aberration-aware training (AAT). Our approach involves a lightweight network that models lens aberrations at different
related reading
- The Little Book of Deep Learningfleuret.org
- Reproducing DeepTFUS | projectsmasonjwang.com
- Foundations of Computer Visionvisionbook.mit.edu
- 3D-Aware Ellipse Prediction for Object-Based Camera Pose Estimationarxiv.org
- Meauring Invariances in Deep Networksai.stanford.edu
- 3D Object Detection via 2D Segmentation-Based Computational Integral Imaging Applied to a Real Videoncbi.nlm.nih.gov
- RAFT: Recurrent All-Pairs Field Transforms for Optical Flow | PDFarxiv.org
- PS$^2$F: Polarized Spiral Point Spread Function for Single-Shot 3D Sensingarxiv.org
- Understanding deep learning requires rethinking generalizationarxiv.org
- Incorporating the image formation process into deep learning improves network performancenature.com
- (Some of) The Models, They Just Don't Want to Learn | Tildeblog.tilderesearch.com
- [2008.05711] Lift, Splat, Shoot: Encoding Images From Arbitrary Camera Rigs by Implicitly Unprojecting to 3Darxiv.org