REVIEW 3 cited by
Learning to Embed Time Series Patches Independently
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Masked time series modeling has recently gained much attention as a self-supervised representation learning strategy for time series. Inspired by masked image modeling in computer vision, recent works first patchify and partially mask out time series, and then train Transformers to capture the dependencies between patches by predicting masked patches from unmasked patches. However, we argue that capturing such patch dependencies might not be an optimal strategy for time series representation learning; rather, learning to embed patches independently results in better time series representations. Specifically, we propose to use 1) the simple patch reconstruction task, which autoencode each patch without looking at other patches, and 2) the simple patch-wise MLP that embeds each patch independently. In addition, we introduce complementary contrastive learning to hierarchically capture adjacent time series information efficiently. Our proposed method improves time series forecasting and classification performance compared to state-of-the-art Transformer-based models, while it is more efficient in terms of the number of parameters and training/inference time. Code is available at this repository: https://github.com/seunghan96/pits.
Forward citations
Cited by 3 Pith papers
-
PreMixer: MLP-Based Pre-training Enhanced MLP-Mixers for Large-scale Traffic Forecasting
A graph-free MLP-Mixer with independent patch-wise MLP masked pretraining matches or beats complex spatiotemporal models on large-scale traffic forecasting at a fraction of the compute.
-
A New Perspective on Time Series Anomaly Detection: Faster Patch-based Broad Learning System
A shallow patch-based broad learning system with a random-perturbation contrastive branch and multi-scale patch ensembling reports state-of-the-art unsupervised time series anomaly detection on five benchmarks.
-
Causal Time-Series Synchronization for Multi-Dimensional Forecasting
Aligning cause-effect pairs by their estimated Granger lag improves channel-dependent forecasting accuracy and transfer learning on synthetic time-series data.
Discussion (0). Continue with ORCID to comment.