BEV-Locator: An End-to-end Visual Semantic Localization Network Using Multi-View Images

11/27/2022
by   Zhihuang Zhang, et al.
0

Accurate localization ability is fundamental in autonomous driving. Traditional visual localization frameworks approach the semantic map-matching problem with geometric models, which rely on complex parameter tuning and thus hinder large-scale deployment. In this paper, we propose BEV-Locator: an end-to-end visual semantic localization neural network using multi-view camera images. Specifically, a visual BEV (Birds-Eye-View) encoder extracts and flattens the multi-view images into BEV space. While the semantic map features are structurally embedded as map queries sequence. Then a cross-model transformer associates the BEV features and semantic map queries. The localization information of ego-car is recursively queried out by cross-attention modules. Finally, the ego pose can be inferred by decoding the transformer outputs. We evaluate the proposed method in large-scale nuScenes and Qcraft datasets. The experimental results show that the BEV-locator is capable to estimate the vehicle poses under versatile scenarios, which effectively associates the cross-model information from multi-view images and global semantic maps. The experiments report satisfactory accuracy with mean absolute errors of 0.052m, 0.135m and 0.251^∘ in lateral, longitudinal translation and heading angle degree.

READ FULL TEXT

page 1

page 7

page 8

page 12

research
07/18/2023

EgoVM: Achieving Precise Ego-Localization using Lightweight Vectorized Maps

Accurate and reliable ego-localization is critical for autonomous drivin...
research
05/07/2023

Poses as Queries: Image-to-LiDAR Map Localization with Transformers

High-precision vehicle localization with commercial setups is a crucial ...
research
04/04/2023

OrienterNet: Visual Localization in 2D Public Maps with Neural Matching

Humans can orient themselves in their 3D environments using simple 2D ma...
research
09/28/2017

X-View: Graph-Based Semantic Multi-View Localization

Global registration of multi-view robot data is a challenging task. Appe...
research
11/11/2022

An Improved End-to-End Multi-Target Tracking Method Based on Transformer Self-Attention

This study proposes an improved end-to-end multi-target tracking algorit...
research
12/19/2022

From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera Calibration

We tackle a new problem of multi-view camera and subject registration in...
research
10/03/2021

Translating Images into Maps

We approach instantaneous mapping, converting images to a top-down view ...

Please sign up or login with your details

Forgot password? Click here to reset