Pith. sign in

Paper Citation Record · LEDGER

Benchmarking and Understanding Compositional Relational Reasoning of LLMs

As of 21 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2412.12841.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12841 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:46:27.679324Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T04:30:42.215749Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T04:30:49.091477Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved4
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 513ffc2f-05b3-47c4-bf87-5e27631667f2 · outbound

This paper cites an unresolved cited work.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:46:27.848653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.649801Z digest=sha256:4d0d56d7716bffcb85f15ec8670e7cfec1a2596c44a625fce51e4bf309e884a7

Observation c4d79150-d5f9-4943-8908-c0b24b483e82 · outbound

This paper cites an unresolved cited work.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:46:27.833801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.654302Z digest=sha256:c9a75cce698addff782468f4153a6bb68c25f456e17b310d1876790ab9e2b9d7

Observation 5e0da18e-023e-4a87-8470-b06ad3e73e1c · outbound

This paper cites vehicle.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs vehicle

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.817687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.658924Z digest=sha256:0fbeb021471f4d56bf1b9011687bed434827b58a277e12efb5b9310e53d92781

Observation b7e3d495-4d6e-445a-b37c-e72cc035341f · outbound

This paper cites affirmative generation) have the same overall structure.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs affirmative generation) have the same overall structure

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.801845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.664680Z digest=sha256:f599094b2c5199ce2b54a7f396d37173eda73b3ebcfef7229a5033cc78243280

Observation d98e33f1-4942-4036-9a1e-583ba0d4be24 · outbound

This paper cites 18.13 PPred.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs 18.13 PPred

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.785854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.669036Z digest=sha256:bc921137178815f04c60b4159b9c3159b9211d570cdd0329bc9d2ea4273ef17a

Observation 4d621e95-61a6-4f02-a468-95a04fe9d129 · outbound

This paper cites heads vary with different rretrieve, because they retrieve different attributes according to the semantic re- lation.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs heads vary with different rretrieve, because they retrieve different attributes according to the semantic re- lation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.770402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.673961Z digest=sha256:61925c5755178022cb95d88418081ca31c66247aa8762f5d51d1112cf21b3929

Observation cbfb198f-3a6f-427b-b75c-96c63349b9c0 · outbound

This paper cites Premise: {premise}. Please answer with Yes or No. Can it be inferred from the premise that {hypothesis}? Answer:→{label}.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs Premise: {premise}. Please answer with Yes or No. Can it be inferred from the premise that {hypothesis}? Answer:→{label}

Reference 13

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T13:46:27.754408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.679324Z digest=sha256:d23213bf92906381c909810a3cc57b8957fe1ab9e687f41fb919b9696b47a439

Observation d50d27c9-7099-40c4-acdf-57357d3f56e8 · outbound

This paper cites In Proceedings of the 2015 Conference on Empir- ical Methods in Natural Language Processing (EMNLP).

Benchmarking and Understanding Compositional Relational Reasoning of LLMs In Proceedings of the 2015 Conference on Empir- ical Methods in Natural Language Processing (EMNLP)

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.912589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.617421Z digest=sha256:5ff6c342f2f8d9de4042bd8b2637c78e629f5489c4d6c06572ebf06993a8c9a4

Observation 4c13c7dc-a374-4afd-9010-aa3f8a0507ba · outbound

This paper cites Transformer Circuits Thread, 1(1): 12.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs Transformer Circuits Thread, 1(1): 12

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.878930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.628024Z digest=sha256:6e98417f8ed11ed062448df1e45856b7144a73b3bcff0c49428760c8f5bff4ff

Observation b2772f64-865c-403c-9d77-92f3862d6589 · outbound

This paper cites In-context Learning and Induction Heads.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs In-context Learning and Induction Heads

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T13:46:27.643661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:46:27.643661Z digest=sha256:1cc68a2691e5917a2550aae15c02762a044e38f5201db995c59f6bf9c9c1e780

Observation a2852e28-5d65-460f-9ff7-49f36ef4f50e · outbound

This paper cites In Conference on Empirical Methods in Natural Language Processing (EMNLP).

Benchmarking and Understanding Compositional Relational Reasoning of LLMs In Conference on Empirical Methods in Natural Language Processing (EMNLP)

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.863491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.632925Z digest=sha256:00786933a1f72161ebb23f2a1e32289e9d3c01e4d9635068242d7378c4d067b8

Observation fde9c21e-bda1-4612-81ed-30fe522c16fe · outbound

This paper cites In Advances in Neural Information Processing Systems (NeurIPS), volume 36.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs In Advances in Neural Information Processing Systems (NeurIPS), volume 36

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:46:27.896109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T13:46:27.622942Z digest=sha256:0b4783cfa99e24249ae98c1dc9c6dadc4d50b1e1ad73199428ba4d4a6a8176f5

Observation 125a6361-79da-4adf-8678-fcfaf5742bf9 · outbound

This paper cites The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets.

Benchmarking and Understanding Compositional Relational Reasoning of LLMs The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets

Reference 2882

Resolution
unresolved
no resolver link, observed 2026-08-11T13:46:27.638083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:46:27.638083Z digest=sha256:9d57b459ea7b0509ff249257eda3551b6e80589341ab8659f9d6a9740d88828b

Pith citing papers

Observation a0bb827e-e8e3-4fb8-9109-87a6f223097e · inbound

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning cites this paper.

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning Benchmarking and Understanding Compositional Relational Reasoning of LLMs

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T04:30:49.095104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T04:30:42.215749Z digest=sha256:88343bf5aae4ed4ed439dfb5f93ea27390b9ddcf1074c9ed19f543210dd0fb09