REVIEW 10 cited by
Towards Time Series Reasoning with LLMs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Multi-modal large language models (MLLMs) have enabled numerous advances in understanding and reasoning in domains like vision, but we have not yet seen this broad success for time-series. Although prior works on time-series MLLMs have shown promising performance in time-series forecasting, very few works show how an LLM could be used for time-series reasoning in natural language. We propose a novel multi-modal time-series LLM approach that learns generalizable information across various domains with powerful zero-shot performance. First, we train a lightweight time-series encoder on top of an LLM to directly extract time-series information. Then, we fine-tune our model with chain-of-thought augmented time-series tasks to encourage the model to generate reasoning paths. We show that our model learns a latent representation that reflects specific time-series features (e.g. slope, frequency), as well as outperforming GPT-4o on a set of zero-shot reasoning tasks on a variety of domains.
Forward citations
Cited by 10 Pith papers
-
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers
A new open-source library and benchmark, xRouteBench, evaluates LLM routers on a shared cost-aware protocol across text, memory, vision, time-series, and personalized tasks.
-
ReasonCast: Towards Explainable Time Series Forecasting with Reasoning
A fine-tuned LLM that states its reasoning, then its forecast, in one response beats specialized forecasters on five synthetic time series patterns.
-
Enhancing LLM Reasoning for Time Series Classification by Tailored Thinking and Fused Decision
A three-turn prompting framework, ReasonTSC, boosts LLM time series classification by fusing pattern analysis with plug-in model scores, but its evaluation leaks test-set labels into the prompts and overstates the gains.
-
T2S: High-resolution Time Series Generation with Text-to-Series Diffusion Models
T2S uses a length-adaptive VAE and flow-matching diffusion transformer to generate variable-length time series from text captions, trained on a new fragment-level caption dataset.
-
CastFSR: A Fast--Slow--Reflect Agentic Reasoning Framework for Context-Aware Time Series Forecasting
CastFSR improves context-aware time series forecasting by combining a fast data-driven forecast prior, slow LLM-driven contextual reasoning, and reflective validation, outperforming most baselines on public benchmarks.
-
MemCast: Memory-Driven Time Series Forecasting with Experience-Conditioned Reasoning
MemCast claims LLM time-series forecasting improves when retrieval from a hierarchical memory of patterns, wisdom, and laws conditions reasoning, but the reported gains depend on a test-label-rewarded confidence update.
-
Causal Graph Fuzzy LLMs: A First Introduction and Applications in Time Series Forecasting
CGF-LLM combines fuzzy time series and PCMCI causal graphs into text input for fine-tuned GPT-2, reporting improved one-step-ahead forecast NRMSE and a reduction in token count on four datasets.
-
TempoGPT: Enhancing Time Series Reasoning via Quantizing Embedding
Converting time series into discrete tokens via a VQ-VAE codebook and feeding them through a shared embedding layer with text improves a language model's accuracy on a synthetic circuit-based time series reasoning benchmark.
-
Human-AI Collaborative Bot Detection in MMORPGs
The paper proposes an unsupervised pipeline of TS2Vec, DBSCAN, and a GPT-4o reviewer for MMORPG bot detection, but validates it only with a proxied access-pattern metric.
-
Large Language models for Time Series Analysis: Techniques, Applications, and Challenges
A review of LLM-based time series analysis that proposes several taxonomies, but is undermined by citation errors and a lack of systematic methodology.
Discussion (0). Continue with ORCID to comment.