Logo image
ViTVO: Vision Transformer based Visual Odometry with Attention Supervision
Conference paper

ViTVO: Vision Transformer based Visual Odometry with Attention Supervision

Chu-Chi Chiu, Hsuan-Kung Yang, Hao-Wei Chen, Yu-Wen Chen and Chun-Yi Lee
Proceedings of MVA 2023 - 18th International Conference on Machine Vision and Applications
2023

Abstract

Artificial Intelligence Computer Graphics and Computer-Aided Design Computer Science Applications Hardware and Architecture
In this paper, we develop a Vision Transformer based visual odometry (VO), called ViTVO. ViTVO introduces an attention mechanism to perform visual odometry. Due to the nature of VO, Transformer based VO models tend to overconcentrate on few points, which may result in a degradation of accuracy. In addition, noises from dynamic objects usually cause difficulties in performing VO tasks. To overcome these issues, we propose an attention loss during training, which utilizes ground truth masks or self supervision to guide the attention maps to focus more on static regions of an image. In our experiments, we demonstrate the superior performance of ViTVO on the Sintel validation set, and validate the effectiveness of our attention supervision mechanism in performing VO tasks.

Metrics

1 Record Views

Details

Logo image