Abstract:Spatial-temporal representation learning is ubiquitous in various real-world applications, including visual comprehension, video understanding, multi-modal analysis, human-computer interaction, and urban computing. Due to the emergence of huge amounts of multi-modal heterogeneous spatial/temporal/spatial-temporal data in big data era, the existing visual methods rely heavily on large-scale data annotations and supervised learning to learn a powerful big model. However, the lack of interpretability, robustness, and out-of-distribution generalization are becoming the bottleneck problems of these models, which hinders the progress of interpretable and reliable artificial intelligence. The majority of the existing methods are based on correlation learning with the assumption that the data are independent and identically distributed, which lack an unified guidance and analysis about why modern spatial-temporal representation learning methods have limited interpretability and easily collapse into dataset bias. Inspired by the strong inference ability of human-level agents, recent years have therefore witnessed great effort in developing causal reasoning paradigms to realize robust representation and model learning with good interpretability. In this paper, we conduct a comprehensive review of existing causal reasoning methods for spatial-temporal representation learning, covering fundamental theories, models, and datasets. The limitations of current methods and datasets are also discussed. Moreover, we propose some primary challenges, opportunities, and future research directions for benchmarking causal reasoning algorithms in spatial-temporal representation learning. This paper aims to provide a comprehensive overview of this emerging field, attract attention, encourage discussions, bring to the forefront the urgency of developing novel causal reasoning methods, publicly available benchmarks, and consensus-building standards for reliable spatial-temporal representation learning and related real-world applications more efficiently.

Deep Causal Inference for Point-referenced Spatial Data with Continuous Treatments

Estimating Direct and Indirect Causal Effects of Spatiotemporal Interventions in Presence of Spatial Interference

Deep treatment-adaptive network for causal inference

Causal-StoNet: Causal Inference for High-Dimensional Complex Data

Causal GNNs: A GNN-Driven Instrumental Variable Approach for Causal Inference in Networks

Semiparametric Regression for Spatial Data via Deep Learning

Causal Inference for Spatial Treatments

DESCN: Deep Entire Space Cross Networks for Individual Treatment Effect Estimation

When causal inference meets deep learning

Causal Reasoning with Spatial-temporal Representation Learning: A Prospective Study

Enhancing the Performance of Neural Networks Through Causal Discovery and Integration of Domain Knowledge

Deep Learning-based Group Causal Inference in Multivariate Time-series

Causal Inference Meets Deep Learning: A Comprehensive Survey

Causal Inference in Spatial Statistics

Causality-Aware Spatiotemporal Graph Neural Networks for Spatiotemporal Time Series Imputation

Nonparametric spatial autoregressive model using deep neural networks

Causal Inference Meets Machine Learning

Spatial causal inference in the presence of unmeasured confounding and interference

Neural Networks with Causal Graph Constraints: A New Approach for Treatment Effects Estimation

Towards Learning and Explaining Indirect Causal Effects in Neural Networks