Neural Reprojection Error: Merging Feature Learning and Camera Pose Estimation

This is the official repository for our paper Neural Reprojection Error: Merging Feature Learning and Camera Pose Estimation , to appear in CVPR 2021. Code to will be released prior to the conference.

Abstract

Absolute camera pose estimation is usually addressed by sequentially solving two distinct subproblems: First a feature matching problem that seeks to establish putative 2D-3D correspondences, and then a Perspective-n-Point problem that minimizes, with respect to the camera pose, the sum of so-called Reprojection Errors (RE). We argue that generating putative 2D-3D correspondences 1) leads to an important loss of information that needs to be compensated as far as possible, within RE, through the choice of a robust loss and the tuning of its hyperparameters and 2) may lead to an RE that conveys erroneous data to the pose estimator. In this paper, we introduce the Neural Reprojection Error (NRE) as a substitute for RE. NRE allows to rethink the camera pose estimation problem by merging it with the feature learning problem, hence leveraging richer information than 2D-3D correspondences and eliminating the need for choosing a robust loss and its hyperparameters. Thus NRE can be used as training loss to learn image descriptors tailored for pose estimation. We also propose a coarse-to-fine optimization method able to very efficiently minimize a sum of NRE terms with respect to the camera pose. We experimentally demonstrate that NRE is a good substitute for RE as it significantly improves both the robustness and the accuracy of the camera pose estimate while being computationally and memory highly efficient. From a broader point of view, we believe this new way of merging deep learning and 3D geometry may be useful in other computer vision applications.

BibTex

Please consider citing our work:

@inproceedings{germain2021NRE,
  author    = {Hugo Germain and
               Vincent Lepetit and
               Guillaume Bourmaud},
  title     = {Neural Reprojection Error: Merging Feature Learning and Camera Pose Estimation},
  booktitle = {CVPR},
  year      = {2021},
  url       = {https://arxiv.org/abs/2103.07153}
}

Neural Reprojection Error: Merging Feature Learning and Camera Pose Estimation

Related tags

Overview

Neural Reprojection Error: Merging Feature Learning and Camera Pose Estimation

Abstract

BibTex

Owner

Hugo Germain

Knowledge Management for Humans using Machine Learning & Tags

Research shows Google collects 20x more data from Android than Apple collects from iOS. Block this non-consensual telemetry using pihole blocklists.

VD-BERT: A Unified Vision and Dialog Transformer with BERT

An original implementation of "MetaICL Learning to Learn In Context" by Sewon Min, Mike Lewis, Luke Zettlemoyer and Hannaneh Hajishirzi

An self sufficient AI that crawls the web to learn how to generate art from keywords

U-Net for GBM

Research code for Arxiv paper "Camera Motion Agnostic 3D Human Pose Estimation"

OptNet: Differentiable Optimization as a Layer in Neural Networks

Platform-agnostic AI Framework 🔥

🧠 A PyTorch implementation of 'Deep CORAL: Correlation Alignment for Deep Domain Adaptation.', ECCV 2016

GPT, but made only out of gMLPs

Text2Art is an AI art generator powered with VQGAN + CLIP and CLIPDrawer models

nnDetection is a self-configuring framework for 3D (volumetric) medical object detection which can be applied to new data sets without manual intervention. It includes guides for 12 data sets that were used to develop and evaluate the performance of the proposed method.

An Open-Source Package for Information Retrieval.

source code and pre-trained/fine-tuned checkpoint for NAACL 2021 paper LightningDOT

[ICCV 2021 Oral] Just Ask: Learning to Answer Questions from Millions of Narrated Videos

hySLAM is a hybrid SLAM/SfM system designed for mapping

FCAF3D: Fully Convolutional Anchor-Free 3D Object Detection

Adaptation through prediction: multisensory active inference torque control

An experimentation and research platform to investigate the interaction of automated agents in an abstract simulated network environments.