Lune

ICCV2019Top-tier venue

Multi-View Stereo by Temporal Nonparametric Fusion

Yuxin Hou, Juho Kannala, Arno Solin

2019Year
99Citations
29Top-tier citations

Abstract

We propose a novel idea for depth estimation from multiview image-pose pairs, where the model has capability to leverage information from previous latent-space encodings of the scene. This model uses pairs of images and poses, which are passed through an encoder-decoder model for disparity estimation. The novelty lies in soft-constraining the bottleneck layer by a nonparametric Gaussian process prior. We propose a pose-kernel structure that encourages similar poses to have resembling latent spaces. The flexibility of the Gaussian process (GP) prior provides adapting memory for fusing information from previous views. We train the encoder-decoder and the GP hyperparameters jointly end-to-end. In addition to a batch method, we derive a lightweight estimation scheme that circumvents standard pitfalls in scaling Gaussian process inference, and demonstrate how our scheme can run in real-time on smart devices.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 6ac783c4-798a-41b1-86c8-3da3e6c7d046

Cited by top-tier papers29

Ask how each one uses it

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines