Pith. sign in

Paper Citation Record · LEDGER

Evaluating Large Language Models for Causal Modeling

As of 13 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2411.15888.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15888 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:50:00.156537Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6f433ea1-dd41-42ad-b6f1-67063f2d4f91 · outbound

This paper cites Causal Parrots: Large Language Models May Talk Causality But Are Not Causal.

Evaluating Large Language Models for Causal Modeling Causal Parrots: Large Language Models May Talk Causality But Are Not Causal

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.104576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.104576Z digest=sha256:f1230ff0a786dcda287661f198b20dbfa9984ecdc522fd120efcac5f475b3749

Observation 925bdd83-16d4-4247-8c5d-0d0d5764be8d · outbound

This paper cites Spock at fincausal 2022: Causal information extraction using span-based and sequence tagging models.

Evaluating Large Language Models for Causal Modeling Spock at fincausal 2022: Causal information extraction using span-based and sequence tagging models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.318619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.120692Z digest=sha256:b036303f8afa46110013ba276e89a77d795d0a89e26d9bd3f85e12aa6b5ca403

Observation 6eb3f997-5420-4fee-892a-5c5c64ca6cff · outbound

This paper cites Causal BERT : Language models for causality detection between events expressed in text.

Evaluating Large Language Models for Causal Modeling Causal BERT : Language models for causality detection between events expressed in text

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.124212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.124212Z digest=sha256:a7f2c75820021d6bca7a64582be0a4a0797821ae6af194bb955f83114944d54b

Observation 4351e96c-2149-4910-8605-0f818e6fba88 · outbound

This paper cites Event causality extraction via implicit cause-effect interactions.

Evaluating Large Language Models for Causal Modeling Event causality extraction via implicit cause-effect interactions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.308736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.127837Z digest=sha256:625a82fb30b8f170df6e2ea5862115a4f8615a14bcbd346e8bb991604af450ee

Observation 2bda8b75-1544-479c-8d47-d1a726d172ad · outbound

This paper cites Causal Reasoning and Large Language Models: Opening a New Frontier for Causality.

Evaluating Large Language Models for Causal Modeling Causal Reasoning and Large Language Models: Opening a New Frontier for Causality

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.131207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.131207Z digest=sha256:1cb139a3f797a8bc5e7e31df7f3de4a374b935d6e0e41b6705095422ed71384a

Observation 2a2f851e-a921-415d-a639-835336fcc362 · outbound

This paper cites Lego: A multi-agent collaborative framework with role-playing and iterative feedback for causality explanation generation.

Evaluating Large Language Models for Causal Modeling Lego: A multi-agent collaborative framework with role-playing and iterative feedback for causality explanation generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.298543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.135173Z digest=sha256:de91141768d9cd8de42d5ccabfd3b40f66ac56b3a2372c9372687f87a1379b9d

Observation bcc87209-af65-4f48-bc9d-3aa223b00d94 · outbound

This paper cites Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis, jan.

Evaluating Large Language Models for Causal Modeling Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis, jan

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.288309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.138738Z digest=sha256:1e12ffc7a710baa6a670d69b764b37f0c09f48ea2c243903426a9ade03a63158

Observation 355dc4bd-7c16-4b26-aff8-66202da1fa6f · outbound

This paper cites The causal reasoning ability of open large language model: A comprehensive and exemplary functional testing.

Evaluating Large Language Models for Causal Modeling The causal reasoning ability of open large language model: A comprehensive and exemplary functional testing

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.278138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.142581Z digest=sha256:18a77043c64274c599b2fe8b2f56f5b553bda015e25f15004e7e302a01db2162

Observation 55dd5eef-6fb2-4840-8cce-bf1240a9aeaa · outbound

This paper cites GPT-4 Technical Report.

Evaluating Large Language Models for Causal Modeling GPT-4 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.145804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.145804Z digest=sha256:36129c0ad77f315587dad550459fffc5fdad3e762e2c00dc9bb5443d71e88446

Observation e57d8f44-46fc-4c5f-883a-a9532d5bff42 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating Large Language Models for Causal Modeling LLaMA: Open and Efficient Foundation Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.149569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.149569Z digest=sha256:91ab154839d09c44f2d3600a8dc8143695bdde3ad010355e2684b31d9681f3ff

Observation 5b800757-cf5f-43a8-9ca5-eae8d7e69f05 · outbound

This paper cites Mixtral of Experts.

Evaluating Large Language Models for Causal Modeling Mixtral of Experts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.153046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.153046Z digest=sha256:c73414ced3d4832bd13a062bce404f9fd9beb8cfa50836e4cc413ec7deb6c584

Observation 737598af-86a8-4751-a4a9-0f2096bec4c6 · outbound

This paper cites Richard Hahn, and Huan Liu.

Evaluating Large Language Models for Causal Modeling Richard Hahn, and Huan Liu

Reference 1960

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.328971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.108583Z digest=sha256:b8c9a65ed9d76def5d599c398cc1201f88293113b94afe4537d650685a996af7

Observation a10dffed-7291-4983-816c-13ac56e94684 · outbound

This paper cites Can Large Language Models Infer Causation from Correlation?.

Evaluating Large Language Models for Causal Modeling Can Large Language Models Infer Causation from Correlation?

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.100482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.100482Z digest=sha256:7623bd275e59050c6c27e58d78c14064bbc651ef3b877798d54c409310ea7865

Observation 3664e219-a616-4056-baed-a5827fb211d0 · outbound

This paper cites doi: 10.1145/3397269.

Evaluating Large Language Models for Causal Modeling doi: 10.1145/3397269

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.112553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.112553Z digest=sha256:8fc97a2139471b9ded7931671ac0175964fc6035bf1ed84d010e68c3179b8051

Observation b7aa4658-a7d4-4021-9d6b-048c8f7a46e9 · outbound

This paper cites TC-GAT: Graph Attention Network for Temporal Causality Discovery.

Evaluating Large Language Models for Causal Modeling TC-GAT: Graph Attention Network for Temporal Causality Discovery

Reference 2022

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:50:00.249584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.116706Z digest=sha256:879a624bde2359c4eeabbe1f366919b180e900a07deb186708011e9172955166

Observation 7f52f8ad-379e-44e1-b9e2-4bf2effdb87b · outbound

This paper cites Weakly supervised multilingual causality extraction from wikipedia.

Evaluating Large Language Models for Causal Modeling Weakly supervised multilingual causality extraction from wikipedia

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.339375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T13:50:00.096055Z digest=sha256:b0b6fc07adea623a3ccd0857155d56192db8c68fb514ced1d076f608156528a4

Observation 4cf55e44-2372-4646-b53c-3693c1926249 · outbound

This paper cites Mistral 7B.

Evaluating Large Language Models for Causal Modeling Mistral 7B

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.156537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.156537Z digest=sha256:aace3331b64df9c9904ff51ffc12dd3e6ff31aae5728161a89b4120e46b1e301

Pith citing papers

No inbound Pith citation observations are available.