arrow
Return

Future object localization using multi-modal ego-centric video

delete2025-12-18
delete0
PRE
AI
J
Jee-Ye Yoon
J
Je‐Won Kang *
DOI:10.1016/j.jvcir.2025.104684delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Future object localization (FOL) seeks to predict the future locations of objects using information from past and present video frames. Ego-centric videos from vehicle-mounted cameras serve as a key source. However, these videos are constrained by a limited field of view and susceptibility to external conditions. To address these challenges, this paper presents a novel FOL approach that combines ego-centric video data with point cloud data, enhancing both robustness and accuracy. The proposed model is based on a deep neural network that prioritizes front-camera ego-centric videos, exploiting their rich visual cues. By integrating point cloud data, the system improves three-dimensional (3D) object localization. Furthermore, the paper introduces a novel method for ego-motion prediction. The ego-motion prediction network employs multi-modal sensors to comprehensively capture physical displacement in both 2D and 3D spaces, effectively handling occlusions and the limited perspective inherent in ego-centric videos. Experimental results indicate that the proposed FOL system with ego-motion prediction (MS-FOLe) outperforms existing methods on large-scale open datasets for intelligent driving.

Journal

Journal of Visual Communication and Image Representation cover
Journal of Visual Communication and Image Representation
IF:
3.1
Papers:
529
Citations:
5.6K

Organization

E
ewha w. university
Scholars:
7
Papers: 3
Citations: 0
Cited Papers

Cited Papers

Realtime Video Latency Reduction for Autonomous Vehicle Teleoperation Using RTMP Over UDP Protocols
err2023-02-27
err0
PREAI
errAna Heryana; Dikdik Krisnandi; Hilman Pardede; Galih Nugraha Nurkahfi; Mochamad Mardi Marta Dinata; Andri Rozie; Rendra Firmansyah
errShare
errSave
BiTraP: Bi-Directional Pedestrian Trajectory Prediction With Multi-Modal Goal Estimation
err2021-04-01
err0
errOAAI
errYu Yao; Ella Atkins; Matthew Johnson-Roberson; Ram Vasudevan; Xiaoxiao Du
errShare
errSave
errShare
errSave
errShare
errSave
Holistic LSTM for Pedestrian Trajectory Prediction
err2021-01-01
err97
PREAI
errQuan, Ruijie; Zhu, Linchao; Wu, Yu; Yang, Yi
errShare
errSave
Multi-Sensor Multi-Vehicle (MSMV) Localization and Mobility Tracking for Autonomous Driving
err2020-12-01
err56
errOAAI
errYang, Pengtao; Duan, Dongliang; Chen, Chen; Cheng, Xiang; Yang, Liuqing
errShare
errSave
M2P3
err2020-03-30
err0
PREAI
errAtanas Poibrenski; Matthias Klusch; Igor Vozniak; Christian Müller
errShare
errSave
researcher View more