Pith. sign in

Paper Citation Record · LEDGER

TRAM: Benchmarking Temporal Reasoning for Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2310.00835.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.00835 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T07:01:42.806076Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:57:41.651539Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6fd0c0e5-5f14-4a5d-ac49-bf8f3cbc4f98 · inbound

Evaluating Very Long-Term Conversational Memory of LLM Agents cites this paper.

Evaluating Very Long-Term Conversational Memory of LLM Agents TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 156

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:05:13.017306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T08:05:10.586357Z digest=sha256:f68fd4a9d702da28a2a57e58931c520e58f36fcd2795d938948aa671d01e2634

Observation 83268337-fb65-4c4e-839e-323f0c56603c · inbound

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models cites this paper.

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T07:01:42.806076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:01:42.806076Z digest=sha256:fb2346fc6e9620e9f0e1b9b95c34633a49b0523056d88e257213e7a51b36736d

Observation ad5c8a61-53be-4996-84b5-1aa5af41b5d2 · inbound

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes cites this paper.

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-21T01:09:20.174294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T01:08:50.547385Z digest=sha256:a306cc9e030e51a1a9aec045dbe6c72f2debbd8b4e4b291037e0d7b6a6d29a24

Observation 80ff6244-8fdd-4fef-9ce4-b19daa0d9be6 · inbound

Reading the Finetuning Prior: Verbatim Content Recovery via Contrastive Decoding Diffing cites this paper.

Reading the Finetuning Prior: Verbatim Content Recovery via Contrastive Decoding Diffing TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T22:44:01.671881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T22:37:02.708284Z digest=sha256:4a8a926604e718ad14677a5c2c8d9fe9b9d05af374e21ce09702c61d64ecba3d

Observation 7d3b205b-ca38-4a71-a10e-819e98590fb1 · inbound

Temporal Preference Concepts and their Functions in a Large Language Model cites this paper.

Temporal Preference Concepts and their Functions in a Large Language Model TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 111

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T14:05:47.192165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T22:16:47.743387Z digest=sha256:6211f67252dbc64efc985a9a53396fc95e94abe4ab37eda0961fc5d5b53db625

Observation 60267c19-8c75-4f70-88ed-8ef36f6c1491 · inbound

Temporal Preference Concepts and their Functions in a Large Language Model cites this paper.

Temporal Preference Concepts and their Functions in a Large Language Model TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 111

Resolution
unresolved
no resolver link, observed 2026-07-12T17:03:44.315006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T17:03:44.315006Z digest=sha256:1beb656ec8e62ce1385006ba3134f80567d4aa5853d5ff3624a897d5b091749b

Observation a91c1ebe-c113-4725-8782-c8a18983701b · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 250

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:57:41.653196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:8a560653c833ca819e4879ebb792cc16c3dc9a3e51154de046beeeed84797582

Observation 7290a414-d2d5-424b-9c5c-5d5fee9b49ef · inbound

WaveformQA: Benchmarking LLM Temporal Reasoning on Digital Waveforms cites this paper.

WaveformQA: Benchmarking LLM Temporal Reasoning on Digital Waveforms TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T09:49:36.498551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:49:36.498551Z digest=sha256:f2b068bd03e2a595a23e6b5b52d808533b6434fab91b2e33f9237cb2567a65af

Observation 625ea4a6-55da-4a5a-be4d-13dc98e7a7a8 · inbound

MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation cites this paper.

MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation TRAM: Benchmarking Temporal Reasoning for Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-31T21:55:17.370287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T21:55:17.370287Z digest=sha256:87f347ccd98233b52a1b39709062e7e7a82799b2be4f6860d8ab2721cc72274b