Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:15:34.146881Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2412.13377.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:15:34.146881Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3639ad6f-dd8e-40a9-a876-24108bfd38f3 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aedaf6f-161c-4481-a29f-e2fb9fcb5c84 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Do All Languages Cost the Same? Tokenization in the Era of Commercial Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fc3bbaa-8027-43ad-bd5e-a5fe445b4bb0 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Interleaving Text and Number Embeddings to Solve Mathemathics Problems
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97527c4e-83c1-4901-a767-2b65809ed6cd · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Language Models are Few-Shot Learners
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c119f560-f1b9-43e7-a5e4-8218a36ca753 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7aee0ba4-1ff6-4d6e-8096-11e1307845a2 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71a21248-58c0-4b15-b128-e53497781a7f · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e3ca96a-2f64-43f2-b1e8-aba675c3728e · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models The Foundations of Tokenization: Statistical and Computational Concerns
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbc92e88-ef04-4bc5-b659-bfcae7782adb · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Unpacking Tokenization: Evaluating Text Compression and its Correlation with Model Performance
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fd843bf-131e-4b3d-b4b1-e20564763615 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models ReTok: Replacing Tokenizer to Enhance Representation Efficiency in Large Language Model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2987062f-3985-42a0-a9cb-056ad6299ac4 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Mistral 7B
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 944cfa6d-db69-4621-8944-3bdc666b68ce · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Unveiling Divergent Inductive Biases of LLMs on Temporal Data
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fa736ab2-de03-4c11-8350-8e09cf6cfe70 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models How Much Can RAG Help the Reasoning of LLM?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ae945cd-9f66-458f-988a-f1c6598b495d · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57d51d96-236e-4787-968d-e1868fe302c0 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models GPT-4o System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fba74f2c-468e-4b91-81bf-1651dd6b1fe1 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models GPT-4 Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 636290a4-b8b9-4fbb-8001-f49995829f29 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a3e9f07-f25e-4dec-a811-0f0467f5ba75 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Toward a Theory of Tokenization in LLMs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f79244-ef34-435c-b3bb-2031095571a3 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Tokenization Is More Than Compression
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05777839-e349-420c-a538-6129381bf7b8 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c155b720-4643-4d42-8301-16958bdbc26d · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5d0477-e022-40ff-a80e-37d32c300c29 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Timo: Towards Better Temporal Reasoning for Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 642a10b6-5547-44ee-bf74-42a85c14a254 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Towards Benchmarking and Improving the Temporal Reasoning Capability of Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e495ff7-ce5e-4682-a39e-176386d8d9ab · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Towards Robust Temporal Reasoning of Large Language Models via a Multi-Hop QA Dataset and Pseudo-Instruction Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b93222fa-ecdb-4722-8f49-f8977a63048f · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7532a3f1-0dff-4fab-9292-29a4748c7b09 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models RedPajama: an Open Dataset for Training Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d11fb6-33d9-4827-abe4-d78f3f42b4ed · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46188d41-63d4-4085-9ad6-7bb571f10584 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Large Language Models Can Learn Temporal Reasoning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04418d5d-aaf7-4fc5-b174-6fa8b5ada241 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Qwen2 Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b58aa1b7-0ef2-4de4-b481-44d4731dba0c · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Counting Ability of Large Language Models and Impact of Tokenization
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3023a51d-8a30-4aac-bea4-2c022352ab68 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Set the Clock: Temporal Alignment of Pretrained Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c708fd36-8ff6-4453-993a-89f6ff0f3e44 · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models Is Your LLM Outdated? A Deep Look at Temporal Generalization
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95d20336-e61f-40df-b686-cb350ea3d46b · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models URL: " 'urlintro :=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e2a6139-a951-448c-aa38-5fb57a2a1efc · outbound
DateLogicQA: Benchmarking Temporal Biases in Large Language Models write newline
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.