arrow
Return

CollaborativeBEV: Collaborative bird eye view for reconstructing crowded environment

delete2024-07-01
delete1
PRE
AI
J
Jiaxin Zhao
F
Fangzhou Mu
Y
Y. F. Lyu *
DOI:10.1016/j.imavis.2024.105060delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Constructing a virtual world for the Metaverse based on real-world data is crucial, yet creating virtual environments for crowded scenes poses challenges in accurately tracking individuals using egocentric wearable cameras due to occlusions caused by crowded pedestrians. To address this, we propose a collaborative perception strategy that leverages multiple agents equipped with multi-view cameras to construct an occupancy map for a crowded environment. To fuse the multi-view perceptions of multiple agents, we propose a Collaborative Bird Eye View fusion network, called CollaborativeBEV (C-BEV), in which, we leverage a depth-based BEV network as a feature extractor, and propose a feature enhancement module to improve perception fusion in overlapping area. A designed loss function is introduced to address data imbalance during training, and a BEV enhancement strategy is proposed to augment the sample pool for training the BEV decoder. Experiment on the Sean2.0 dataset demonstrates that our C-BEV method performs better than the baseline method in terms of a 5.3% IoU increase. Our code will be released on github. https://github.com/RYaNzzZ1/CollaborativeBEV.
Keywords:
Metaverse
3D reconstruction
Occupancy map
Collaborative BEV
Point cloud understanding

Journal

Image and Vision Computing cover
Image and Vision Computing
IF:
4.2
Papers:
4.1K
Citations:
6.7K

Organization

S
southeast university - china
Scholars:
5.3W
Papers: 4.9W
Citations: 57
Cited Papers

Cited Papers

Multi-Agent Systems: A Survey
err2018-01-01
err506
errOAAI
errDorri, Ali; Kanhere, Salil S.; Jurdak, Raja
errShare
errSave
SGF3D: Similarity-guided fusion network for 3D object detection
err2024-02-01
err4
PREAI
errLi, Chunzheng; Wang, Gaihua; Long, Qian; Zhou, Zhengshu
errShare
errSave
A survey on deep multimodal learning for computer vision: advances, trends, applications, and datasets
err2021-06-10
err192
errOAAI
errBayoudh, Khaled; Knani, Raja; Hamdaoui, Faycal; Mtibaa, Abdellatif
errShare
errSave
err
IF0
err
err0
PREAI
err
errShare
errSave
RGB-D salient object detection: A survey
err2021-03-01
err205
errOAAI
errZhou, Tao; Fan, Deng-Ping; Cheng, Ming-Ming; Shen, Jianbing; Shao, Ling
errShare
errSave
3D-VDNet: Exploiting the vertical distribution characteristics of point clouds for 3D object detection and augmentation
err2022-11-01
err5
PREAI
errXiao, Weiping; Li, Xiaomao; Liu, Chang; Gao, Jiantao; Luo, Jun; Peng, Yan; Zhou, Yang
errShare
errSave
researcher View more