arrow
Return

Causal reasoning in typical computer vision tasks

delete2023-12-14
delete0
delete
OA
AI
K
KeXuan Zhang
Q
Qiyu Sun
C
Chaoqiang Zhao
Y
Yang Tang *
DOI:10.1007/s11431-023-2502-9delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Deep learning has revolutionized the field of artificial intelligence. Based on the statistical correlations uncovered by deep learning-based methods, computer vision tasks, such as autonomous driving and robotics, are growing rapidly. Despite being the basis of deep learning, such correlation strongly depends on the distribution of the original data and is susceptible to uncontrolled factors. Without the guidance of prior knowledge, statistical correlations alone cannot correctly reflect the essential causal relations and may even introduce spurious correlations. As a result, researchers are now trying to enhance deep learning-based methods with causal theory. Causal theory can model the intrinsic causal structure unaffected by data bias and effectively avoids spurious correlations. This paper aims to comprehensively review the existing causal methods in typical vision and vision-language tasks such as semantic segmentation, object detection, and image captioning. The advantages of causality and the approaches for building causal paradigms will be summarized. Future roadmaps are also proposed, including facilitating the development of causal theory and its application in other complex scenarios and systems.
Keywords:
causal reasoning
computer vision tasks
vision-language tasks
semantic segmentation
object detection

Journal

Science China-Technological Sciences cover
Science China-Technological Sciences
IF:
4.9
Papers:
4.9K
Citations:
9.9K

Organization

No organization information available