arrow
Return

SpatialScene2Vec: A self-supervised contrastive representation learning method for spatial scene similarity evaluation

delete2024-04-01
delete2
delete
OA
AI
D
Danhuai Guo *
Y
Yingxue Yu
S
Song Gao
G
Gengchen Mai
H
Huixuan Chen
DOI:10.1016/j.jag.2024.103743delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Spatial scene similarity plays a crucial role in spatial cognition, as it enables us to understand and compare different spatial scenes and their relationships. However, understanding spatial scenes is a complex task. While existing literature has contributed to spatial scene representation learning, these methods primarily focus on comprehending the spatial relationships among objects, often neglecting their semantic features. Furthermore, there is a lack of scene representation learning methods that can seamlessly handle different types of spatial objects (e.g., points, polylines, and polygons) in a scene. Moreover, since expert knowledge is required for the annotation process of spatial scene understanding, publicly available high-quality annotation data has a limited size which usually leads to suboptimal results. To address these issues, we propose a novel multiscale spatial scene encoding model called SpatialScene2Vec. SpatialScene2Vec utilizes a point location encoder to seamlessly encode the spatial information of different types of spatial objects. A point feature encoder is employed to encode the semantic features of these objects. A spatial scene embedding is generated by integrating the spatial embeddings and feature embeddings of spatial objects within this scene. Furthermore, to address the limited labeled data problem, we propose a self-supervised learning framework to train the SpatialScene2Vec model in which a contrastive loss is used for spatial scene similarity evaluation. In addition, we introduce a novel spatial scene data augmentation method to generate positive scene augmentations by leveraging the unique characteristics of spatial scenes and random sampling points based on the shapes of polyline/polygon objects within the current spatial scenes. We conduct experiments on real -world datasets for spatial scene retrieval tasks, including vector data types of points, polylines, and polygons. Results show that SpatialScene2Vec outperforms well-established encoding methods such as Space2Vec due to the advantages of the integrated multi-scale representations and the proposed spatial scene data augmentation method, with significant improvements and robustness.
Keywords:
Contrastive learning
Representation learning
Spatial scene
Similarity assessment
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

International Journal of Applied Earth Observation and Geoinformation cover
International Journal of Applied Earth Observation and Geoinformation
IF:
8.6
Papers:
5.1K
Citations:
2.4W

Organization

U
university of wisconsin madison
Scholars:
3.8W
Papers: 2.9W
Citations: 53
L
lenovo
Scholars:
131
Papers: 110
Citations: 1
University of Wisconsin System cover
University of Wisconsin System
Scholars:
6.7W
Papers: 5.8W
Citations: 382
B
Beijing University of Chemical Technology
Scholars:
3.1W
Papers: 2.2W
Citations: 4.5W
L
legend holdings
Scholars:
208
Papers: 177
Citations: 1
researcher View more organizations