Pith. sign in

REVIEW 8 cited by

Large Language Models can Deliver Accurate and Interpretable Time Series Anomaly Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.15370 v1 pith:HITAN6WK submitted 2024-05-24 cs.CL

classification cs.CL
keywords detectionllmadtsadanomalyllmsemploysmodelsseries
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Time series anomaly detection (TSAD) plays a crucial role in various industries by identifying atypical patterns that deviate from standard trends, thereby maintaining system integrity and enabling prompt response measures. Traditional TSAD models, which often rely on deep learning, require extensive training data and operate as black boxes, lacking interpretability for detected anomalies. To address these challenges, we propose LLMAD, a novel TSAD method that employs Large Language Models (LLMs) to deliver accurate and interpretable TSAD results. LLMAD innovatively applies LLMs for in-context anomaly detection by retrieving both positive and negative similar time series segments, significantly enhancing LLMs' effectiveness. Furthermore, LLMAD employs the Anomaly Detection Chain-of-Thought (AnoCoT) approach to mimic expert logic for its decision-making process. This method further enhances its performance and enables LLMAD to provide explanations for their detections through versatile perspectives, which are particularly important for user decision-making. Experiments on three datasets indicate that our LLMAD achieves detection performance comparable to state-of-the-art deep learning methods while offering remarkable interpretability for detections. To the best of our knowledge, this is the first work that directly employs LLMs for TSAD.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 8 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Argos: Agentic Time-Series Anomaly Detection with Autonomous Rule Generation via Large Language Models

    cs.LG 2025-01 reject novelty 7.0 of 10

    ARGOS uses LLM agents to generate explainable, reproducible anomaly detection rules and fuses them with a base detector, reporting higher F1 than deep-learning and LLM baselines on KPI, Yahoo, and a Microsoft internal...

  2. LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers

    cs.CL 2026-08 conditional novelty 6.0 of 10

    A new open-source library and benchmark, xRouteBench, evaluates LLM routers on a shared cost-aware protocol across text, memory, vision, time-series, and personalized tasks.

  3. C-RE-ACT: Causal RE-ACTing Agent for O-RAN Forensic Triage

    cs.NI 2026-07 reject novelty 6.0 of 10

    An agentic O-RAN triage system that ranks root causes via SAM causal discovery and graph soft-prompting claims 89% top-3 accuracy on 140 testbed experiments.

  4. UFO2: The Desktop AgentOS

    cs.AI 2025-04 reject novelty 6.0 of 10

    UFO2 reports that deep Windows integration, hybrid vision and UIA control detection, and GUI-plus-API actions lift desktop automation success rates above prior CUAs, but the evaluation is partly contaminated by benchm...

  5. Large Action Models: From Inception to Implementation

    cs.AI 2024-12 conditional novelty 6.0 of 10

    A four-phase training pipeline converts a 7B language model into a Windows GUI action model that reaches 81.2% offline and 71.0% online task success on the authors' Word test set, beating text-only GPT-4o.

  6. Human-AI Collaborative Bot Detection in MMORPGs

    cs.AI 2025-08 reject novelty 4.0 of 10

    The paper proposes an unsupervised pipeline of TS2Vec, DBSCAN, and a GPT-4o reviewer for MMORPG bot detection, but validates it only with a proxied access-pattern metric.

  7. Interpretable Anomaly-Based DDoS Detection in AI-RAN with XAI and LLMs

    cs.CR 2025-07 conditional novelty 4.0 of 10

    An LSTM trained on 5G key performance measurements detects DDoS attacks from user equipment with F1 above 0.96, while LIME, SHAP, and an LLM convert decisions into human-readable explanations and suggested mitigations.

  8. Foundation Models for Anomaly Detection: Vision and Challenges

    cs.LG 2025-02 conditional novelty 4.0 of 10

    A survey that taxonomizes foundation-model-based anomaly detection into encoder, detector, and interpreter roles and lists open challenges.

Pith tools