REVIEW 19 cited by
Large Language Models for Time Series: A Survey
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large Language Models (LLMs) have seen significant use in domains such as natural language processing and computer vision. Going beyond text, image and graphics, LLMs present a significant potential for analysis of time series data, benefiting domains such as climate, IoT, healthcare, traffic, audio and finance. This survey paper provides an in-depth exploration and a detailed taxonomy of the various methodologies employed to harness the power of LLMs for time series analysis. We address the inherent challenge of bridging the gap between LLMs' original text data training and the numerical nature of time series data, and explore strategies for transferring and distilling knowledge from LLMs to numerical time series analysis. We detail various methodologies, including (1) direct prompting of LLMs, (2) time series quantization, (3) aligning techniques, (4) utilization of the vision modality as a bridging mechanism, and (5) the combination of LLMs with tools. Additionally, this survey offers a comprehensive overview of the existing multimodal time series and text datasets and delves into the challenges and future opportunities of this emerging field. We maintain an up-to-date Github repository which includes all the papers and datasets discussed in the survey.
Forward citations
Cited by 19 Pith papers
-
Efficient and Adaptive Human Activity Recognition via LLM Backbones
Pretrained LLMs adapted via convolutional projections and LoRA act as efficient frozen backbones for sensor-based human activity recognition, delivering strong data efficiency and cross-dataset transfer.
-
TS-Agent: Understanding and Reasoning Over Raw Time Series via Iterative Insight Gathering
TS-Agent is an agentic framework that uses LLMs only for evidence-based reasoning while delegating extraction to raw time series tools, matching or exceeding baselines on four benchmarks with largest gains on reasoning tasks.
-
Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training
An RL agent learns domain re-weighting policies from evaluation feedback to improve balanced performance in continual pre-training of LLMs across source and target domains.
-
A Cost-Effective Multimodal LLM Reasoning Framework for Question Answering over Irregular Clinical Time Series
ClinPRISM reaches 49.83% average accuracy on CLIR-Bench irregular clinical time-series QA using a 4B LLM, 16 temporal tokens, and 0.15 s/question.
-
CLIR-Bench: Benchmarking Multimodal Question Answering over Irregular Clinical Time Series
CLIR-Bench shows generalist and time-series LLMs struggle to ground clinical answers in sparse irregular ICU evidence, with top accuracy near 50% and weak causal evidence use.
-
CausalMoE: A Billion-Scale Multimodal Foundation Model for Granger Causal Discovery with Pattern-Routed Heterogeneous Experts
CausalMoE is a multimodal foundation model with pattern-routed heterogeneous experts and LLM/VLM integration that claims new SOTA performance on supervised and few-shot Granger causal discovery benchmarks.
-
InA-Probe: Instruction-Aware Active Probing for Time Series Forecasting with LLMs
InA-Probe improves LLM time series forecasting via instruction-aware active probing, outperforming baselines with up to 37% error reduction on seven benchmarks in one-for-all and zero-shot settings.
-
Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference
Two-stage LLM framework infers stellar parameters and ~20 elemental abundances from spectra, showing performance gains with increasing data volume.
-
Time Series Augmented Generation for Financial Applications
TSAG lets LLMs use external tools for financial time series analysis, with a new benchmark showing capable agents achieve near-perfect tool accuracy and minimal hallucination.
-
Natural Language Interfaces for Spatial and Temporal Databases: A Comprehensive Overview of Methods, Taxonomy, and Future Directions
A literature survey that taxonomizes methods, datasets, and evaluation practices for natural language interfaces to geospatial and temporal databases while identifying recurring trends and future directions.
-
TSAQA: Time Series Analysis Question And Answering Benchmark
TSAQA provides 210k QA samples across 13 domains and six tasks, showing current LLMs score at most 65.08% zero-shot and struggle most with temporal-order reasoning.
-
BEDTime: A Unified Benchmark for Automatically Describing Time Series
BEDTime benchmark tests 17 models on describing time series structure and finds vision-language models outperform dedicated time-series-language models and language-only approaches, with all models fragile to robustne...
-
Time Series Forecasting as Reasoning: A Slow-Thinking Approach with Reinforced LLMs
Time-R1 trains LLMs via supervised fine-tuning followed by reinforcement learning with a time-series-specific reward and non-uniform GRIP sampling to enable multi-step reasoning that improves forecasting accuracy.
-
Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference
A two-stage LLM framework infers stellar parameters and ~20 elemental abundances from spectra, with performance improving systematically as training data volume increases.
-
Spectra as Language: Large Language Models for Scalable Stellar Parameter and Abundance Inference
A two-stage LLM framework infers stellar parameters and ~20 elemental abundances from spectra, with performance improving as training data increases.
-
Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning
Supervised fine-tuning of pretrained LLMs on offline trajectories yields better few-shot sequential decision-making than in-context-only baselines, with a theoretical suboptimality bound derived for linear MDPs by int...
-
Physics-Aware LLM-Based Probabilistic Wind Power Scenario Generation under Extreme Icing Conditions
A physics-aware LLM framework generates high-fidelity probabilistic wind power scenarios under extreme icing by enforcing physical constraints like power limits and ramp rates on trajectories from real SCADA data.
-
ELATE: Evolutionary Language model for Automated Time-series Engineering
An LLM-guided evolutionary feature engineering method for time-series forecasting reduces RMSE by 8.4% on average across seven datasets.
-
Technology-assisted Personalized Yoga for Better Health -- Challenges and Outlook
A vision statement for personalized yoga decision support whose claimed content cannot be verified, because the supplied full text is a different paper on maritime drift prediction.
Discussion (0). Sign in to comment.