Pith. sign in

Paper Citation Record · LEDGER

Evaluating Large Language Models for Causal Modeling

As of 13 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2411.15888.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15888 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:50:00.156537Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6f433ea1-dd41-42ad-b6f1-67063f2d4f91 · outbound

This paper cites Causal Parrots: Large Language Models May Talk Causality But Are Not Causal.

Evaluating Large Language Models for Causal Modeling Causal Parrots: Large Language Models May Talk Causality But Are Not Causal

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.104576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.104576Z digest=sha256:74ded3a1113872c4e48df5642636c845556a03e197ce0bd5132652c8ebcb6d20

Observation 925bdd83-16d4-4247-8c5d-0d0d5764be8d · outbound

This paper cites Spock at fincausal 2022: Causal information extraction using span-based and sequence tagging models.

Evaluating Large Language Models for Causal Modeling Spock at fincausal 2022: Causal information extraction using span-based and sequence tagging models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.318619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.120692Z digest=sha256:c6b1c08d469fbe7f147eaf157c5e620520f51adaaf7e287e7219662c21bf22a2

Observation 6eb3f997-5420-4fee-892a-5c5c64ca6cff · outbound

This paper cites Causal BERT : Language models for causality detection between events expressed in text.

Evaluating Large Language Models for Causal Modeling Causal BERT : Language models for causality detection between events expressed in text

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.124212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.124212Z digest=sha256:a7f2c75820021d6bca7a64582be0a4a0797821ae6af194bb955f83114944d54b

Observation 4351e96c-2149-4910-8605-0f818e6fba88 · outbound

This paper cites Event causality extraction via implicit cause-effect interactions.

Evaluating Large Language Models for Causal Modeling Event causality extraction via implicit cause-effect interactions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.308736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.127837Z digest=sha256:7a81443316c56a86dac9425f7c12475f0db111e4dcc7b7f0591e58f6fa6652cc

Observation 2bda8b75-1544-479c-8d47-d1a726d172ad · outbound

This paper cites Causal Reasoning and Large Language Models: Opening a New Frontier for Causality.

Evaluating Large Language Models for Causal Modeling Causal Reasoning and Large Language Models: Opening a New Frontier for Causality

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.131207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.131207Z digest=sha256:1cb139a3f797a8bc5e7e31df7f3de4a374b935d6e0e41b6705095422ed71384a

Observation 2a2f851e-a921-415d-a639-835336fcc362 · outbound

This paper cites Lego: A multi-agent collaborative framework with role-playing and iterative feedback for causality explanation generation.

Evaluating Large Language Models for Causal Modeling Lego: A multi-agent collaborative framework with role-playing and iterative feedback for causality explanation generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.298543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.135173Z digest=sha256:4c88dab6bda9a194a90465a1f6c2124b1736dbccf9d053f8de49cd2b1b93afdb

Observation bcc87209-af65-4f48-bc9d-3aa223b00d94 · outbound

This paper cites Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis, jan.

Evaluating Large Language Models for Causal Modeling Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis, jan

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.288309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.138738Z digest=sha256:1d52fbd310875819c1a7f25aa6f09f0b96c0b9ebca4790bc74f0f24c9096dc34

Observation 355dc4bd-7c16-4b26-aff8-66202da1fa6f · outbound

This paper cites The causal reasoning ability of open large language model: A comprehensive and exemplary functional testing.

Evaluating Large Language Models for Causal Modeling The causal reasoning ability of open large language model: A comprehensive and exemplary functional testing

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.278138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.142581Z digest=sha256:a54f81751c0cc6882c2e478f0fe159e7b6c968773f461ce87b96211d561b5c39

Observation 55dd5eef-6fb2-4840-8cce-bf1240a9aeaa · outbound

This paper cites GPT-4 Technical Report.

Evaluating Large Language Models for Causal Modeling GPT-4 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.145804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.145804Z digest=sha256:36129c0ad77f315587dad550459fffc5fdad3e762e2c00dc9bb5443d71e88446

Observation e57d8f44-46fc-4c5f-883a-a9532d5bff42 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating Large Language Models for Causal Modeling LLaMA: Open and Efficient Foundation Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.149569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.149569Z digest=sha256:91ab154839d09c44f2d3600a8dc8143695bdde3ad010355e2684b31d9681f3ff

Observation 5b800757-cf5f-43a8-9ca5-eae8d7e69f05 · outbound

This paper cites Mixtral of Experts.

Evaluating Large Language Models for Causal Modeling Mixtral of Experts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.153046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.153046Z digest=sha256:c73414ced3d4832bd13a062bce404f9fd9beb8cfa50836e4cc413ec7deb6c584

Observation 737598af-86a8-4751-a4a9-0f2096bec4c6 · outbound

This paper cites Richard Hahn, and Huan Liu.

Evaluating Large Language Models for Causal Modeling Richard Hahn, and Huan Liu

Reference 1960

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.328971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.108583Z digest=sha256:065a494b7a10fcbd6138045d9368e0bde4233652df1042963e28044924f7a6a0

Observation a10dffed-7291-4983-816c-13ac56e94684 · outbound

This paper cites Can Large Language Models Infer Causation from Correlation?.

Evaluating Large Language Models for Causal Modeling Can Large Language Models Infer Causation from Correlation?

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.100482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.100482Z digest=sha256:bb13bcf84d1ae71bef5b698d178e2baa5109f5adec9553b67a896e3d4df7ed9f

Observation 3664e219-a616-4056-baed-a5827fb211d0 · outbound

This paper cites doi: 10.1145/3397269.

Evaluating Large Language Models for Causal Modeling doi: 10.1145/3397269

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.112553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.112553Z digest=sha256:8fc97a2139471b9ded7931671ac0175964fc6035bf1ed84d010e68c3179b8051

Observation b7aa4658-a7d4-4021-9d6b-048c8f7a46e9 · outbound

This paper cites TC-GAT: Graph Attention Network for Temporal Causality Discovery.

Evaluating Large Language Models for Causal Modeling TC-GAT: Graph Attention Network for Temporal Causality Discovery

Reference 2022

Resolution
verified exact
local_arxiv, observed 2026-08-12T13:50:00.249584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.116706Z digest=sha256:096d341406a98c9f389fcd55bb2e1eafd347af5532d62df264782c3ac9ea7a8a

Observation 7f52f8ad-379e-44e1-b9e2-4bf2effdb87b · outbound

This paper cites Weakly supervised multilingual causality extraction from wikipedia.

Evaluating Large Language Models for Causal Modeling Weakly supervised multilingual causality extraction from wikipedia

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T13:50:00.339375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T13:50:00.096055Z digest=sha256:bbe3c17e4d5ab40b1bc5881ba011b3141ec57edb57e48562f9e6772236cff687

Observation 4cf55e44-2372-4646-b53c-3693c1926249 · outbound

This paper cites Mistral 7B.

Evaluating Large Language Models for Causal Modeling Mistral 7B

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T13:50:00.156537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:50:00.156537Z digest=sha256:aace3331b64df9c9904ff51ffc12dd3e6ff31aae5728161a89b4120e46b1e301

Pith citing papers

No inbound Pith citation observations are available.