Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:12:43.381851Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2412.17970.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:12:43.381851Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4e321e14-b6e4-4359-9ae2-ae5a99b319c6 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Which modality should I use - text, motif, or image? : Understanding graphs with large language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cce26841-3931-42dd-a016-4695e79722ef · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models GPT4Graph: Can Large Language Models Understand Graph Structured Data ? An Empirical Evaluation and Benchmarking
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddbb3e12-ee5a-4089-9ef8-8fe37ba65873 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7abd8f6-3351-44fa-8e1f-cfe5ef9951e2 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Can Long-Context Language Models Subsume Retrieval, RAG, SQL, and More?
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efe915b9-89b2-45f7-a9c2-a6c82c355041 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models CLEVR-Math: A Dataset for Compositional Language, Visual and Mathematical Reasoning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 410715b7-b360-46f4-8897-74b558352d82 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ec14147-348a-42af-a786-60faa50d9ce3 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e205a35-27de-4950-8d71-61de680c88c3 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models CELLO: Causal Evaluation of Large Vision-Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db5fc4f9-a60e-4e3f-8d4b-8368119f7a73 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models doi: 10.1162/tacl a 00446
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbdea996-a866-4be3-bbbd-a7927e5fec55 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 882aeca8-3b24-46d1-bc10-8f1058cb3e44 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models LLM Processes: Numerical Predictive Distributions Conditioned on Natural Language
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03a001b8-ab05-49cf-8762-f2a6f9e3f65c · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Causal Evaluation of Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9d36e33-eac4-4069-b3f4-806d1e71fc7b · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models URL https: //www.kaggle.com/m/3301
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cc21b9c-a814-47ff-8612-ebe2951d6a21 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a74b270-bf06-45eb-afd1-f0f17112696a · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models GITA: Graph to Visual and Textual Integration for Vision-Language Graph Reasoning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8517cd0-0800-479a-abca-ef29594ac560 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models doi: 10.3115/v1/P15-1142
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b50c1fa0-d44f-413b-9fca-8e79188792fe · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Mistral 7B
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44acbef8-94bc-49bf-896c-66c3d8321ee3 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Training Verifiers to Solve Math Word Problems
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1388eede-7ba2-46c9-989c-07deafe7eef6 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models doi: 10.18653/v1/2020.emnlp-main.89
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 839995f1-f280-4122-8ade-a246f0db9a3c · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Qwen Technical Report
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6152efc-e87d-4250-ac75-cc54b726fc34 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Evaluating Large Language Models on Graphs: Performance Insights and Comparative Analysis
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac627b00-9b4d-48eb-baa4-a5df380691f4 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models The Essential Role of Causality in Foundation World Models for Embodied AI
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54801bf9-0f58-4a8c-87cf-6bfa13526d57 · outbound
CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.