Return
Temporal consistent multi-view perception for robust embodied manipulation
DOI:10.1016/j.patcog.2025.112177.png)
Abstract
En 中文
• A two-stage framework combining contrastive and multi-view imitation for manipulation. • Align visual representations with instructions to enhance temporal consistency. • TMVP significantly boosts few-shot learning, outperforming multi-task baselines.
Keywords:
contrastive learning
multi-view imitation
manipulation
few-shot learning
temporal consistency
Journal
IF:
7.6
Papers:
1.3W
Citations:
4.5W

