PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibility

Lee, Jae Yong; DeGol, Joseph; Zou, Chuhang; Hoiem, Derek

Computer Science > Computer Vision and Pattern Recognition

arXiv:2108.08943 (cs)

[Submitted on 19 Aug 2021]

Title:PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibility

Authors:Jae Yong Lee, Joseph DeGol, Chuhang Zou, Derek Hoiem

View PDF

Abstract:Recent learning-based multi-view stereo (MVS) methods show excellent performance with dense cameras and small depth ranges. However, non-learning based approaches still outperform for scenes with large depth ranges and sparser wide-baseline views, in part due to their PatchMatch optimization over pixelwise estimates of depth, normals, and visibility. In this paper, we propose an end-to-end trainable PatchMatch-based MVS approach that combines advantages of trainable costs and regularizations with pixelwise estimates. To overcome the challenge of the non-differentiable PatchMatch optimization that involves iterative sampling and hard decisions, we use reinforcement learning to minimize expected photometric cost and maximize likelihood of ground truth depth and normals. We incorporate normal estimation by using dilated patch kernels, and propose a recurrent cost regularization that applies beyond frontal plane-sweep algorithms to our pixelwise depth/normal estimates. We evaluate our method on widely used MVS benchmarks, ETH3D and Tanks and Temples (TnT), and compare to other state of the art learning based MVS models. On ETH3D, our method outperforms other recent learning-based approaches and performs comparably on advanced TnT.

Comments:	Accepted to ICCV 2021 for oral presentation
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2108.08943 [cs.CV]
	(or arXiv:2108.08943v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2108.08943

Submission history

From: Jae Yong Lee [view email]
[v1] Thu, 19 Aug 2021 23:14:48 UTC (24,851 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibility

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibility

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators