Top 4 papers published by David G. Lowe from University of British Columbia in 2017

Proceedings Article•DOI•

Unsupervised Learning of Depth and Ego-Motion from Video

[...]

Tinghui Zhou¹, Matthew Brown², Noah Snavely², David G. Lowe²•Institutions (2)

University of California, Berkeley¹, Google²

25 Apr 2017

TL;DR: In this paper, an unsupervised learning framework for the task of monocular depth and camera motion estimation from unstructured video sequences is presented, which uses single-view depth and multiview pose networks with a loss based on warping nearby views to the target using the computed depth and pose.

...read moreread less

Abstract: We present an unsupervised learning framework for the task of monocular depth and camera motion estimation from unstructured video sequences. In common with recent work [10, 14, 16], we use an end-to-end learning approach with view synthesis as the supervisory signal. In contrast to the previous work, our method is completely unsupervised, requiring only monocular video sequences for training. Our method uses single-view depth and multiview pose networks, with a loss based on warping nearby views to the target using the computed depth and pose. The networks are thus coupled by the loss during training, but can be applied independently at test time. Empirical evaluation on the KITTI dataset demonstrates the effectiveness of our approach: 1) monocular depth performs comparably with supervised methods that use either ground-truth pose or depth for training, and 2) pose estimation performs favorably compared to established SLAM systems under comparable input settings.

...read moreread less

1,972 citations

Posted Content•

Unsupervised Learning of Depth and Ego-Motion from Video

[...]

Tinghui Zhou¹, Matthew Brown², Noah Snavely², David G. Lowe²•Institutions (2)

University of California, Berkeley¹, Google²

25 Apr 2017-arXiv: Computer Vision and Pattern Recognition

TL;DR: Empirical evaluation demonstrates the effectiveness of the unsupervised learning framework for monocular depth performs comparably with supervised methods that use either ground-truth pose or depth for training, and pose estimation performs favorably compared to established SLAM systems under comparable input settings.

...read moreread less

Abstract: We present an unsupervised learning framework for the task of monocular depth and camera motion estimation from unstructured video sequences. We achieve this by simultaneously training depth and camera pose estimation networks using the task of view synthesis as the supervisory signal. The networks are thus coupled via the view synthesis objective during training, but can be applied independently at test time. Empirical evaluation on the KITTI dataset demonstrates the effectiveness of our approach: 1) monocular depth performing comparably with supervised methods that use either ground-truth pose or depth for training, and 2) pose estimation performing favorably with established SLAM systems under comparable input settings.

...read moreread less

717 citations

Posted Content•

Learning with Imprinted Weights

[...]

Hang Qi, Matthew Brown, David G. Lowe

19 Dec 2017

TL;DR: The process weight imprinting is called as it directly sets weights for a new category based on an appropriately scaled copy of the embedding layer activations for that training example, which provides immediate good classification performance and an initialization for any further fine-tuning in the future.

...read moreread less

Abstract: Human vision is able to immediately recognize novel visual categories after seeing just one or a few training examples. We describe how to add a similar capability to ConvNet classifiers by directly setting the final layer weights from novel training examples during low-shot learning. We call this process weight imprinting as it directly sets weights for a new category based on an appropriately scaled copy of the embedding layer activations for that training example. The imprinting process provides a valuable complement to training with stochastic gradient descent, as it provides immediate good classification performance and an initialization for any further fine-tuning in the future. We show how this imprinting process is related to proxy-based embeddings. However, it differs in that only a single imprinted weight vector is learned for each novel category, rather than relying on a nearest-neighbor distance to training instances as typically used with embedding methods. Our experiments show that using averaging of imprinted weights provides better generalization than using nearest-neighbor instance embeddings.

...read moreread less

16 citations

Posted Content•

Low-Shot Learning with Imprinted Weights

[...]

Hang Qi¹, Matthew Brown², David G. Lowe²•Institutions (2)

University of California, Los Angeles¹, Google²

19 Dec 2017-arXiv: Computer Vision and Pattern Recognition

TL;DR: Weight imprinting as discussed by the authors directly sets weights for a new category based on an appropriately scaled copy of the embedding layer activations for that training example, which provides immediate good classification performance and initialization for any further fine-tuning in the future.

...read moreread less

Abstract: Human vision is able to immediately recognize novel visual categories after seeing just one or a few training examples. We describe how to add a similar capability to ConvNet classifiers by directly setting the final layer weights from novel training examples during low-shot learning. We call this process weight imprinting as it directly sets weights for a new category based on an appropriately scaled copy of the embedding layer activations for that training example. The imprinting process provides a valuable complement to training with stochastic gradient descent, as it provides immediate good classification performance and an initialization for any further fine-tuning in the future. We show how this imprinting process is related to proxy-based embeddings. However, it differs in that only a single imprinted weight vector is learned for each novel category, rather than relying on a nearest-neighbor distance to training instances as typically used with embedding methods. Our experiments show that using averaging of imprinted weights provides better generalization than using nearest-neighbor instance embeddings.

...read moreread less

11 citations

Showing papers by "David G. Lowe published in 2017"