Pith. sign in

Paper Citation Record · LEDGER

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning

As of 5 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2602.14200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.14200 v6

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T23:21:33.792875Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f6e35d26-f5a5-4663-9990-cb3a8eacbf8d · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:32.032061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:32.032061Z digest=sha256:5bf04d01c49889dd18794a8d211de2a7b258b2e1f12770fb7ebb5897f6a26911

Observation ff08d5b8-8386-4c1c-bcda-32aa23461964 · outbound

This paper cites Time-LLM: Time Series Forecasting by Reprogramming Large Language Models.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Time-LLM: Time Series Forecasting by Reprogramming Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:32.114578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:32.114578Z digest=sha256:9ce482c5c328215cc5ec33b2684a1c621d25b39924241a0902339571f83fe42e

Observation a90d04dd-6e7f-45ff-89ba-3efef0ffa91e · outbound

This paper cites OpenTSLM: Time-series language models for reasoning over multivariate medical text- and time-series data.arXiv preprint arXiv:2510.02410,.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning OpenTSLM: Time-series language models for reasoning over multivariate medical text- and time-series data.arXiv preprint arXiv:2510.02410,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:32.317144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:32.317144Z digest=sha256:43d5d71c520265067d0746c901a9a6cf82f467d1860e9c19183a28afc46651ff

Observation 6ebfac6f-2adf-43c1-b359-ce24dc68ec14 · outbound

This paper cites Answer: <your answer>.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Answer: <your answer>

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.792875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.792875Z digest=sha256:5934bff32d31e703c2e69b67213cde7b3c653d6c0f8a997279585ca848d0d8d8

Observation 68050259-8d4e-433c-a576-9a40e7355167 · outbound

This paper cites These annotations support training and evaluation of reasoning capabilities beyond direct answer extraction.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning These annotations support training and evaluation of reasoning capabilities beyond direct answer extraction

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.668082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.668082Z digest=sha256:8ce041ce5d4798f684e23fb048709d1aa27ab3fb6745339a42733f84ba5b7301

Observation d376a4f4-3614-4ad4-acb0-2f5ddf8c3785 · outbound

This paper cites Their benchmark evaluates five tasks: forecasting, classification, anomaly detection, and imputation, with emphasis on limited supervision settings.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Their benchmark evaluates five tasks: forecasting, classification, anomaly detection, and imputation, with emphasis on limited supervision settings

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.286639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.286639Z digest=sha256:0f04592fbd292338eb69c786fc756a2d9bcf9b5cacede853cce019ce4b47cd38

Observation 91b75c10-876d-4bdf-a5c5-ae9af584d15f · outbound

This paper cites Their TSQA dataset comprises approximately 200k question-answer pairs across diverse domains.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Their TSQA dataset comprises approximately 200k question-answer pairs across diverse domains

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.388406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.388406Z digest=sha256:508a917eee7ce08ce97449815421856b8a834dd6bc9c1b1057365b65afb60be6

Observation 853ea64f-86cf-49c6-8be3-d9b5b46b7a58 · outbound

This paper cites 02:34:56:789 AM.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning 02:34:56:789 AM

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.531172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.531172Z digest=sha256:808f765cbbda873c989e4da32a79520ea31bd6ba94f3e0bb4b8aec07f1c5ca8a

Observation 2ee1a13d-da36-4745-9eb0-174ef41ddc0b · outbound

This paper cites Random” indicates uniform subsampling to the budget cap; “full.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Random” indicates uniform subsampling to the budget cap; “full

Reference 151

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.177141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.177141Z digest=sha256:c8389b4ebb78390246c44ad76118c6dfdd9c78f6fd1301c06c1c1ecd98b9166b

Observation da0c4c6b-93ac-49ec-9e0e-efd8479606f3 · outbound

This paper cites TimeSeriesExam: A time series understanding exam.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning TimeSeriesExam: A time series understanding exam

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:31.709054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:31.709054Z digest=sha256:439267154d28c01321bfd68dd3454e8dab2524a06e3cf6dd2c30f9091c82674e

Observation 94e88e0f-9735-4b2b-bfab-e730f6bd0fdf · outbound

This paper cites Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:32.745298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:32.745298Z digest=sha256:7676d318cfd24e989c90fa38dddb943c5061807c3c700e3a1db40fad9c88d969

Observation 82a771da-cb53-4e2b-9eb7-bcc6bda9766e · outbound

This paper cites A CLASSIFICATIONEXPERIMENTDETAILS A.1 EXPERIMENTALSETUP Model architecture.We use the Flamingo variant of OpenTSLM (Langer et al.,.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning A CLASSIFICATIONEXPERIMENTDETAILS A.1 EXPERIMENTALSETUP Model architecture.We use the Flamingo variant of OpenTSLM (Langer et al.,

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.047086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:33.047086Z digest=sha256:f1664f60606830d07615666ae2d70eef1a4f06036d67f856582625d2fc5768f9

Observation bbc2b0ce-27ad-4083-84ad-9f945449bf44 · outbound

This paper cites GPT-4 Technical Report.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:32.570892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:32.570892Z digest=sha256:58d1a5cc1c7cac971e766a39214037eeedb8cc4b8489f7baa15c4f825ef310dc

Observation f49f33e6-ef1c-4899-b652-8b762eddd68a · outbound

This paper cites The Llama 3 Herd of Models.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning The Llama 3 Herd of Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:31.882313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:31.882313Z digest=sha256:30339a176b9706da51bbc42b4d97921bc960b61b1ef9ebf2062b0232ed3735d0

Observation 71647806-738d-43da-bbdd-549690ea3a6f · outbound

This paper cites ITFormer: Bridging Time Series and Natural Language for Multi-Modal QA with Large-Scale Multitask Dataset.

TS-Haystack: A Multi-Task Retrieval Benchmark for Long-Context Time-Series Reasoning ITFormer: Bridging Time Series and Natural Language for Multi-Modal QA with Large-Scale Multitask Dataset

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:32.958348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:21:32.958348Z digest=sha256:06b8385cc0b340543a7e6d5ea5fe9bc66f9613bd57e56ca74b132603ae528c0a

Pith citing papers

No inbound Pith citation observations are available.