Skip to content

Cross-city few-shot spatiotemporal graph forecasting via masked pre-training and prompt tuning.

Aug 2026 · Neural Networks · Vol 205 Pt C, pp. 109553 · 0 citations · 47 references
Medicine

TL;DR

A novel STG few-shot learning framework named ST-MPPT is proposed, which addresses both challenges through masked pre-training and prompt tuning, and introduces a novel prompt network that dynamically generates input-specific prompts to steer the pre-trained encoders to adapt to different data distributions across diverse cities.

Abstract

Spatiotemporal Graph (STG) forecasting holds great significance in the field of urban computing. However, the challenge of data scarcity poses significant obstacles to this task. While cross-city few-shot learning offers a promising solution, existing methods face two fundamental challenges: 1) insufficient extraction of meta-knowledge from data-rich source cities, and 2) limited generality of the knowledge transfer mechanism. In this paper, we propose a novel STG few-shot learning framework named ST-MPPT, which addresses both challenges through masked pre-training and prompt tuning. In the pre-training stage, we perform spatiotemporal-decoupled masked pre-training on source cities with abundant data, enabling the model to learn long-term spatiotemporal patterns more comprehensively. In the downstream forecasting stage, we leverage the pre-trained encoders to acquire robust spatial and temporal representations. These representations are then used to construct a graph structure and enhance the downstream spatiotemporal predictor. To achieve a more general knowledge transfer, we introduce a novel prompt network. Instead of rigid pattern retrieval, this network dynamically generates input-specific prompts to steer the pre-trained encoders to adapt to different data distributions across diverse cities. Extensive experiments on four real-world spatiotemporal datasets demonstrate the superiority of ST-MPPT over strong and representative baselines.

View source

Similar papers

2026

STNet: Multi-Scale Spatiotemporal Learning and Adaptive Fusion for Few-Shot Tor Traffic Classification

The Tor network’s anonymity is increasingly exploited for cybercrime, creating a demand for accurate traffic classification under strict few-shot constraints. While recent efforts like WF-Transformer demonstrate strong temporal modeling capabilities, they still require abundant labeled data and struggle to generalize u...

De-Peng Xu, Guo-Zhen Cheng, Hong-Chao Hu et al. · 0 citations
Conference Open access Sep 2026

ASTPKEFormer: Adaptive Spatiotemporal Prior Knowledge Embedding-Induced Transformers for Traffic Data Forecasting

ASTPKEformer is proposed, a prior knowledge-guided Trans-former framework for traffic prediction that outperforms state-of-the-art (SOTA) baselines, validating its effectiveness and enhancing the representation capability.

Wenfeng Zhou, Xiao-Yun Xia, Xiang-Jie Kong et al. · 0 citations
Book Open access Aug 2026

Orbit-Adaptive Zero-Shot Forecasting on Spatio-Temporal Graph

This work proposes \sysname, an Orbit-Adaptive Graph Neural Network framework, which outperforms existing transfer and naive zero-shot baselines, establishing a new paradigm for structure-driven zero-shot forecasting.

Yue Xu, Wen-Ying Duan, Xiao-Xi He et al. · 0 citations
Book Open access Aug 2026

Where to Go Next: Enhancing Zero-Shot Capability for Cross-City Mobility Prediction

ReLoX is a framework for enhancing zero-shot capability for cross-city mobility prediction that combines an anchor-centered relative trajectory representation with a local egocentric image that captures nearby spatial relations and semantics and achieves substantial gains under cross-city zero-shot protocols, where cit...

Tianao Sun, Kai Zhao, Weiming Huang et al. · 0 citations
Conference Open access Sep 2026

Two-Stage Fine-Grained Trajectory Generation Constrained by Road Networks

Trajectory generation is a pivotal technique for mitigating data sparsity, but existing methods struggle to simultaneously achieve strict road network alignment and capture realistic movement characteristics. To bridge this gap, we propose RNTrajGen, a two-stage fine-grained trajectory generation framework constrained...

Ze-Wu Lv, Zi-Pei Fan, Zhiwen Zhang et al. · 0 citations
Preprint Aug 2026

Rethinking Pre-Training and Augmentation for Zero-Shot Cross-City Object Detection

Real-world deployment of traffic surveillance systems is bottlenecked by geographic domain shift, in which models trained in one city underperform when applied to an unseen target city. Conventional domain adaptation relies on hyperparameter-sensitive architectures or direct profiling of target data. Both are fundament...

Long Hoang Pham, Quoc Pham-Nam Ho, Huy-Hung Nguyen et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.