Logo image
Multi-Modal Pedestrian Crossing Intention Prediction with Transformer-Based Model
期刊文章   開放取用(OA)   同儕審查

Multi-Modal Pedestrian Crossing Intention Prediction with Transformer-Based Model

Ting-Wei Wang尚宏 賴
APSIPA Transactions on Signal and Information Processing, 卷.13(5)
2024

摘要

Autonomous driving;Multi-modal learning

Pedestrian crossing intention prediction based on computer vision plays a pivotal role in enhancing the safety of autonomous driving and advanced driver assistance systems. In this paper, we present a novel multi-modal pedestrian crossing intention prediction framework leveraging the transformer model. By integrating diverse sources of information and leveraging the transformer’s sequential modeling and parallelization capabilities, our system accurately predicts pedestrian crossing intentions. We introduce a novel representation of traffic environment data and incorporate lifted 3D human pose and head orientation data to enhance the model’s understanding of pedestrian behavior. Experimental results demonstrate the state-of-the-art accuracy of our proposed system on benchmark datasets.

檔案與連結 (1)

url
https://doi.org/10.1561/116.20240019檢視
已出版(紀錄版本) 開放

相關連結

指標

1 檢視次數

詳細資料

Logo image