Hyperspectral Trajectory Image for Multi-Month Trajectory Anomaly Detection
cs.CV, cs.LG
Submitted: 2026-03-26
Updated: 2026-09-28
License: http://creativecommons.org/licenses/by/4.0/
The gist: Trajectory anomaly detection underpins applications from fraud detection to urban mobility analysis.
Terminology
Abstract
Trajectory anomaly detection underpins applications from fraud detection to urban mobility analysis. Dense GPS methods preserve fine-grained evidence such as abnormal speeds and short-duration events, but their quadratic cost makes multi-month analysis intractable; consequently, no existing approach detects anomalies over multi-month dense GPS trajectories. The field instead relies on scalable sparse stay-point methods that discard this evidence, forcing separate architectures for each regime and preventing knowledge transfer. Dense and sparse trajectories share a two-dimensional cyclic structure along within-day and across-day axes. We therefore propose TITAnD (Trajectory Image Transformer for Anomaly Detection), which reformulates trajectory anomaly detection as a vision problem by representing trajectories as a Hyperspectral Trajectory Image (HTI): a day x time-of-day grid whose channels encode spatial, semantic, temporal, and kinematic information from either modality, unifying both under a single representation. Under this formulation, agent-level detection reduces to image classification and temporal localization to semantic segmentation. To model this representation, we introduce the Cyclic Factorized Transformer (CFT), which factorizes attention along the two temporal axes, encoding the cyclic inductive bias of human routines, while reducing attention cost by orders of magnitude and enabling dense multi-month anomaly detection for the first time. Empirically, TITAnD achieves the best AUC-PR across the three primary sparse and dense benchmarks, surpassing vision models like UNet while being 11-75x faster than the Transformer with comparable memory, demonstrating that vision reformulation and structure-aware modeling are jointly essential. Additional evaluation includes PoL, LM-TAD, Chronos-2, and DINOv2+Adapter; TITAnD trails LM-TAD on PoL. Code will be made public soon.
Sources
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- Axial Attention in Multidimensional Transformers
- TrajMamba: An Efficient and Semantic-rich Vehicle Trajectory Pre-training Model
- TimeMixer++: A General Time Series Pattern Machine for Universal Predictive Analysis
- Uncertainty-aware Human Mobility Modeling and Anomaly Detection
- UniTraj: Learning a Universal Trajectory Foundation Model from Billion-Scale Worldwide Traces
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models