REVIEW 36 cited by
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
read the original abstract
Over the past years, foundation models have caused a paradigm shift in machine learning due to their unprecedented capabilities for zero-shot and few-shot generalization. However, despite the success of foundation models in modalities such as natural language processing and computer vision, the development of foundation models for time series forecasting has lagged behind. We present Lag-Llama, a general-purpose foundation model for univariate probabilistic time series forecasting based on a decoder-only transformer architecture that uses lags as covariates. Lag-Llama is pretrained on a large corpus of diverse time series data from several domains, and demonstrates strong zero-shot generalization capabilities compared to a wide range of forecasting models on downstream datasets across domains. Moreover, when fine-tuned on relatively small fractions of such previously unseen datasets, Lag-Llama achieves state-of-the-art performance, outperforming prior deep learning approaches, emerging as the best general-purpose model on average. Lag-Llama serves as a strong contender to the current state-of-art in time series forecasting and paves the way for future advancements in foundation models tailored to time series data.
Forward citations
Cited by 36 Pith papers
-
FactoryNet: A Large-Scale Dataset toward Industrial Time-Series Foundation Models
FactoryNet is the first universal pretraining corpus for industrial time-series data with a shared S-E-F-C schema that supports cross-embodiment transfer and competitive anomaly detection.
-
Expert-Guided Forecast Editing for Time-Series Foundation Models
DEFT edits frozen time-series foundation-model forecasts by exploiting model samples and searching over trend/seasonal components, improving forecast quality under small expert-query budgets.
-
UC-Search: Risk-Aware Test-Time Search for Delayed Constrained Time-Series Control
UC-Search is a model-agnostic test-time wrapper that adds feasibility-automaton search and uncertainty-based risk adjustment to produce better delayed constrained control than CEM, MPPI, and risk-random baselines on p...
-
GlucoFM: A Dual-Stream Foundation Model for Continuous Glucose Monitoring
GlucoFM decomposes CGM traces into dual state-event streams, pretrains on 109k hours of unlabeled data, and reports superior subject-disjoint performance on seven clinical tasks across four cohorts.
-
SurF: A Generative Model for Multivariate Irregular Time Series Forecasting
SurF applies the Time Rescaling Theorem as a learnable bijection to create a single generative model for forecasting irregular multivariate event streams that outperforms or matches baselines on six benchmarks.
-
TimeClaw: A Time-Series AI Agent with Exploratory Execution Learning
TimeClaw is an exploratory execution learning system that turns multiple valid tool-use paths into hierarchical distilled experience for improved time-series reasoning without test-time adaptation.
-
FactoryNet: A Large-Scale Dataset toward Industrial Time-Series Foundation Models
FactoryNet is a 51M-point industrial time-series dataset with an S-E-F-C schema that supports zero-shot cross-embodiment transfer and competitive anomaly detection across robotic and machining tasks.
-
Explainable Load Forecasting with Covariate-Informed Time Series Foundation Models
Time series foundation models match the performance of specialized models for day-ahead load forecasting while providing explanations that match domain knowledge on weather and calendar effects.
-
TempusBench: An Evaluation Framework for Time-Series Forecasting
TempusBench is a new evaluation framework for time-series forecasting models that supplies fresh non-overlapping datasets, tasks beyond horizon and domain, consistent tuning across models, and visualization tools.
-
Sundial: A Family of Highly Capable Time Series Foundation Models
Sundial uses TimeFlow Loss for native pre-training of Transformers on continuous time series from TimeBench, achieving SOTA point and probabilistic forecasting with millisecond inference.
-
Deep Time Series Models: A Comprehensive Survey and Benchmark
This survey and benchmark of deep time series models using the released TSLib library finds that models with specific structures perform well only on distinct analysis tasks.
-
CENTILE: A Telemetry Foundation Model Evaluated by the Decisions It Drives
One pretrained telemetry model, CENTILE, improves both HPC backfilling and ISP capacity provisioning decisions under replay, with zero-shot transfer across months and domains.
-
LLM as Forecasting Planner: Training-Free Text Conditioning for Time-Series Foundation Models
LAFP uses MCTS so a frozen TSFM proposes forecast pieces and an LLM ranks and judges them against text, improving text-conditioned forecasts without retraining.
-
LiFT-MPC: Language-in-the-Loop Feedback Tuning of Cost Previews for MPC
Language-informed residual corrections of cost previews, tuned online by a delayed control-performance loss, improve MPC economic performance with a proven regret-style bound.
-
Residual-Guided Multi-Resolution Refinement of Foundation Models: A Case Study in Drought Forecasting
A residual-guided, coarse-to-fine inference wrapper consistently improves frozen time-series foundation models for monthly drought-index forecasting, cutting one-month-ahead MSE by up to 18.9%.
-
A Benchmark for Electrical Load Forecasting Across Grid Levels: Time-Series Transformers Outperform Established Methods
Transformers—especially a standard encoder-decoder—yield the lowest hourly load forecast errors across TSO, low-voltage feeder, and client-level datasets, with 6.6–10.7% error reduction over the best non-Transformer baseline.
-
When Do Foundation Models Pay Off? A Break-Even Analysis of Pretrained Time Series Forecasters
Time-series foundation models are unconditionally better than classical methods on 15/30 datasets, lose early on 6, and a n_train<700 + seasonality rule resolves 10 deployment decisions without training.
-
Trend strength predicts when generative foundation models win: a power-controlled benchmark, a mechanism, and an actionable selection rule
Zero-shot Chronos wins time-series benchmarks by under-extrapolating trend, and trend strength computed before forecasting predicts when it will beat classical models.
-
Beyond Tokenization: Direct Timestep Embedding and Contrastive Alignment for Time-Series Question Answering
CADE framework uses direct timestep embedding and supervised contrastive alignment to improve time-series question answering, reporting gains on six tasks in the Time-MQA benchmark over LLM baselines.
-
TimeRouter: Efficient and Adaptive Routing of Time-Series Foundation Models
TimeRouter routes among time-series foundation models via discriminative routing, selective gating and ensemble fallback, reporting SOTA LB MASE 0.6765 on GIFT-EVAL.
-
Conditional Imputation for Within-Modality Missingness in Multi-Modal Federated Learning
CondI applies conditional diffusion models in a two-phase federated pipeline to impute within-modality missing data, then trains extractors on the completed inputs for downstream tasks on clinical datasets.
-
Time-Aware Prior Fitted Networks for Zero-Shot Forecasting with Exogenous Variables
ApolloPFN trains a time-aware prior-data fitted network on synthetic time series with exogenous variables and outperforms existing zero-shot forecasters on M5 and electricity price benchmarks.
-
DeXposure-FM: A Time-series, Graph Foundation Model for Credit Exposures and Stability on Decentralized Financial Networks
A fine-tuned GraphPFN foundation model forecasts DeFi exposure networks and stress-test losses, beating learned baselines everywhere and persistence on link statistics, though not on average stress-test error.
-
Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?
Time-series foundation models are better calibrated than ARIMA and N-BEATS baselines on the tested datasets, and their calibration does not show the systematic overconfidence seen in image and language models.
-
Hopformer: Homogeneity-Pursuit Transformer for Time Series Forecasting
A two-stage forecaster (SPA trend extraction + LoRA-fine-tuned residual Transformer) that the paper claims beats prior models by 6.56% MASE, though the claim is not robust to its own extended baseline tables.
-
Lightweight Wrappers for Adapting Time Series Foundation Models to Regional Drought Forecasting
Inference-time wrappers that add multi-resolution residual corrections or block-bootstrap averaging to frozen time-series foundation models reduce MSE for regional one-month-ahead SPEI forecasts.
-
VAIOM: Continuous-Input, Discrete-Output Decoder-Only Financial Sequence Modeling
A continuous-input, categorical-output decoder-only Transformer improves held-out one-hour FX return likelihood over single-bar LightGBM and statistical baselines.
-
Zeus: Towards Tuning-Free Foundation Model for Time Series Analysis
Zeus proposes a multi-scale Transformer with point-wise tokenization and Multi-Objective Temporal Masking to enable tuning-free performance on forecasting, interpolation, and other time series tasks.
-
UC-Search: Risk-Aware Test-Time Search for Delayed Constrained Time-Series Control
Bounded retained search over frozen time-series traces improves delayed constrained first actions only under delayed feasible-set coupling, retained-prefix margins, and fail-closed release certificates, with one promo...
-
Benchmarking Deep Time Series Models for Equity Portfolios
Benchmark of 15 time-series architectures on equity portfolios finds no model dominates, with TransEnc-8 at 0.352 rank-1 acceptability and all promoted models showing negative net Sharpe at 20 bps costs under constraints.
-
PaP-NF: Probabilistic Long-Term Time Series Forecasting via Prefix-as-Prompt Reprogramming and Normalizing Flows
PaP-NF uses prefix-as-prompt reprogramming of a frozen LLM to extract global context that conditions a normalizing flow decoder, producing probabilistic long-term time series forecasts evaluated by CRPS.
-
Heterogeneous Scientific Foundation Model Collaboration
Eywa enables language-based agentic AI systems to collaborate with specialized scientific foundation models for improved performance on structured data tasks.
-
Thermal-GEMs: Generalized Models for Building Thermal Dynamics
Multi-source transfer learning for building thermal dynamics yields up to 63% lower forecasting errors than single-source models and outperforms time series foundation models when pretrained on 16-32 buildings over one year.
-
Contextual Deconvolution for Variance-Stable Demand Sensing: Kernel-Modulated Operators in Promotional Retail
A smooth-baseline-plus-sparse-shock decomposition lowers forecast variance and safety stock but increases stockout costs, reducing total inventory cost only when holding costs exceed ~20% of stockout costs.
-
From Vector Autoregressions to AI-based Time Series Forecasting: A Review
AI forecasting methods are flexible generalizations of the classical VAR's conditional forecast distribution, gaining adaptability and scale but losing ready-made inference, identification, and structural interpretation.
-
Time Series Analysis in Machine Learning
A review chapter covering basic time series concepts, classical models like ARIMA, and ML approaches including neural networks and transformers.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.