Skip to content
Open access

Physics-Informed Generative Modeling for Sparse OD Matrix Forecasting: A Dual-Head VAE-GMM Approach With a Single Time Step Input

2026 · IEEE Access · Vol 14, pp. 123707-123726 · 0 citations · 65 references
Computer Science

TL;DR

This study proposes a novel, compact, physics-informed generative framework capable of forecasting transit demands using only a single historical time step as input, which successfully mitigates zero-inflation and performs significantly better than traditional parametric and deep-learning baselines.

Abstract

Accurate passenger flow prediction is crucial to the development of efficient public transit systems, enabling optimal resource allocation, dynamic route planning, and robust congestion mitigation. However, forecasting highly sparse origin-destination (OD) matrices remains a persistent challenge due to extreme data overdispersion and the “zero convergence” problem, where models gravitate toward predicting zeros to minimize global error at the expense of edge-level accuracy. Traditional sequential forecasting models often exacerbate these issues by relying on extensive historical data sequences, which compounds memory overhead and limits real-time scalability. In order to address these limitations, this study proposes a novel, compact, physics-informed generative framework (the Dual-Head VAE-GMM) capable of forecasting transit demands using only a single historical time step as input. The proposed architecture incorporates a Gaussian Mixture Model (GMM) latent prior to capture multi-modal mobility regimes and utilizes a dual-head decoder that explicitly decouples binary network topology edge prediction from continuous volume regression. To ensure predictions align with real-world flow dynamics, the model is governed by a composite objective function that integrates physics-informed graph spectral regularizers and domain-aware mass conservation laws into a generalized variational lower bound. Evaluations on a comprehensive smart card dataset from the 748-station Seoul metropolitan subway network indicate that the proposed approach successfully mitigates zero-inflation and performs significantly better than traditional parametric and deep-learning baselines. The model achieves a mean absolute error of under two passengers with an inference latency of about 57 milliseconds for a 60-minute forecasting horizon.

Read PDF

Similar papers

Conference Open access Sep 2026

PROB-EMOE: A Probabilistic Ensemble Mixture-of-Experts Framework for Metro Network Expansion Forecasting

Forecasting Origin-Destination (OD) demand for new metro lines is critical for sustainable infrastructure planning but faces spatiotemporal out-of-distribution challenges. Existing models often struggle to capture heterogeneous interaction patterns in changing topologies and overlook inherent uncertainty and over-dispe...

Fang-Yi Ding, Zhan Zhao, Zhi Li et al. · 0 citations
Preprint Aug 2026

A Multi-View Coupled Tensor Decomposition for Lightweight Online Adaptive Traffic Prediction

Accurate online traffic prediction is essential for intelligent transportation systems, where forecasting must be performed continuously under imperfect sensing conditions. Missing observations and anomalous disturbances make this task challenging, particularly when prediction relies on a single traffic view. This pape...

Quan Yu, Jie Ni, Yu-Hong Dai et al. · 0 citations
Open access Aug 2026

GSPINN: A Graph Sequential Physics-Informed Surrogate for Trip Travel Time Prediction

This work shows that embedding physically meaningful structure into learning objectives is an effective strategy for traffic surrogate modeling, yielding models that maintain competitive predictive accuracy while substantially improving directional behavioral consistency.

Blessing Itoro Afolayan, Arka Ghosh, Santhanakrishnan Narayanan et al. · 0 citations

Persistent Structure Meets Dynamic Attention: Cross-Variable Priors for Multivariate Time Series Forecasting

A Params-Per-Pair diagnostic is introduced that predicts from dataset properties alone whether structural priors will help and reveals a horizon-dependent complementarity: the structural prior contributes 33% of the gain at short horizons but 88% at long horizons, confirming that time-invariant knowledge compensates as...

C. Mohapatra, Rohit Malshe, J. Pachón · 0 citations
#machine learning Preprint Aug 2026

M3-Former: Multimodal Transformer with Mixture-of-Experts for Long-Term Vessel Trajectory Prediction

To address the challenges of behavioral multimodality, limited semantic utilization, and long-term error accumulation in vessel trajectory prediction, this paper proposes M3-Former, a multimodal trajectory prediction framework enhanced by large language models (LLMs). The proposed framework incorporates vessel static a...

Wen-Long Jin, Hai-Na Tang · 0 citations
#machine learning Preprint Aug 2026

Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

A training-free method for multi-modal trajectory prediction that achieves comparable accuracy to a 57M-parameter transformer while requiring no GPU and zero learned parameters, which enables deployment in new geographic regions from an order of magnitude less historical data.

Michael Fore, Akshay Jain, J. Downes et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.