Leveraging Deep Visual Descriptors for Hierarchical Efficient Localization

Sarlin, Paul-Edouard; Debraine, Frédéric; Dymczyk, Marcin; Siegwart, Roland; Cadena, Cesar

Computer Science > Computer Vision and Pattern Recognition

arXiv:1809.01019 (cs)

[Submitted on 4 Sep 2018 (v1), last revised 18 Sep 2018 (this version, v2)]

Title:Leveraging Deep Visual Descriptors for Hierarchical Efficient Localization

Authors:Paul-Edouard Sarlin, Frédéric Debraine, Marcin Dymczyk, Roland Siegwart, Cesar Cadena

View PDF

Abstract:Many robotics applications require precise pose estimates despite operating in large and changing environments. This can be addressed by visual localization, using a pre-computed 3D model of the surroundings. The pose estimation then amounts to finding correspondences between 2D keypoints in a query image and 3D points in the model using local descriptors. However, computational power is often limited on robotic platforms, making this task challenging in large-scale environments. Binary feature descriptors significantly speed up this 2D-3D matching, and have become popular in the robotics community, but also strongly impair the robustness to perceptual aliasing and changes in viewpoint, illumination and scene structure. In this work, we propose to leverage recent advances in deep learning to perform an efficient hierarchical localization. We first localize at the map level using learned image-wide global descriptors, and subsequently estimate a precise pose from 2D-3D matches computed in the candidate places only. This restricts the local search and thus allows to efficiently exploit powerful non-binary descriptors usually dismissed on resource-constrained devices. Our approach results in state-of-the-art localization performance while running in real-time on a popular mobile platform, enabling new prospects for robotics research.

Comments:	CoRL 2018 Camera-ready (fix typos and update citations)
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1809.01019 [cs.CV]
	(or arXiv:1809.01019v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1809.01019

Submission history

From: Paul-Edouard Sarlin [view email]
[v1] Tue, 4 Sep 2018 14:25:17 UTC (7,477 KB)
[v2] Tue, 18 Sep 2018 20:51:28 UTC (7,549 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Leveraging Deep Visual Descriptors for Hierarchical Efficient Localization

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Leveraging Deep Visual Descriptors for Hierarchical Efficient Localization

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators