Pith. sign in

REVIEW 5 cited by

Scaleformer: Iterative Multi-scale Refining Transformers for Time Series Forecasting

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2206.04038 v4 pith:OTST63AW submitted 2022-06-08 cs.LG

classification cs.LG
keywords seriestimeforecastingacrossarchitecturedatasetsdemonstrateimprovements
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The performance of time series forecasting has recently been greatly improved by the introduction of transformers. In this paper, we propose a general multi-scale framework that can be applied to the state-of-the-art transformer-based time series forecasting models (FEDformer, Autoformer, etc.). By iteratively refining a forecasted time series at multiple scales with shared weights, introducing architecture adaptations, and a specially-designed normalization scheme, we are able to achieve significant performance improvements, from 5.5% to 38.5% across datasets and transformer architectures, with minimal additional computational overhead. Via detailed ablation studies, we demonstrate the effectiveness of each of our contributions across the architecture and methodology. Furthermore, our experiments on various public datasets demonstrate that the proposed improvements outperform their corresponding baseline counterparts. Our code is publicly available in https://github.com/BorealisAI/scaleformer.

Discussion (0). Sign in to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. LeapTS: Rethinking Time Series Forecasting as Adaptive Multi-Horizon Scheduling

    cs.LG 2026-05 unverdicted novelty 7.0 of 10

    LeapTS reformulates forecasting as adaptive multi-horizon scheduling via hierarchical control and NCDEs, delivering at least 7.4% better performance and 2.6-5.3x faster inference than Transformer baselines while adapt...

  2. Amortized Predictability-aware Training Framework for Time Series Forecasting and Classification

    cs.LG 2026-02 conditional novelty 6.0 of 10

    APTF reweights training samples by loss-based predictability buckets and uses an amortization model to stabilize the estimates, improving accuracy across TSF and TSC baselines.

  3. The Power of Architecture: Deep Dive into Transformer Architectures for Long-Term Time Series Forecasting

    cs.LG 2025-07 conditional novelty 6.0 of 10

    Bidirectional joint-attention, complete forecasting aggregation, and direct mapping form the most effective Transformer design for long-term time series forecasting.

  4. SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

    cs.LG 2026-02 conditional novelty 4.0 of 10

    SEMixer combines random-mask patch interactions with progressive adjacent-scale mixing and reports improved MSE/MAE on common long-term forecasting benchmarks.

  5. Foundation Models for Clean Energy Forecasting: A Comprehensive Review

    eess.SY 2025-07 conditional novelty 3.0 of 10

    A survey of foundation model methods, data, and open problems for renewable energy forecasting, built from roughly 218 cited works.

Pith tools