Pith. sign in

Paper Citation Record · LEDGER

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures

As of 9 August 2026, this Paper Citation Record lists 9 of 9 outbound references and 0 inbound Pith citation observations for arXiv:2505.24069.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24069 v4

Coverage vector

measured 9 of 9 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:42:39.596354Z

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

9 of 9 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved5
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c931fb64-0f01-4a99-a450-5eff50ef3507 · outbound

This paper cites Mixtral of Experts.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Mixtral of Experts

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:38.862158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:38.862158Z digest=sha256:547b784e99babeae938db294ce6ce1a35f44ed0e254743ecda49e6aa3020bf05

Observation f9b46fef-c915-48e2-8cf3-c2d7b4a95500 · outbound

This paper cites Code Simulation Challenges for Large Language Models.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Code Simulation Challenges for Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.128432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.128432Z digest=sha256:76626df722a5b0086783e9946ee3bfc4a3094c20d97171dbdb9492cfc1936612

Observation 520f27ae-e1ca-41b4-9077-2fdfa797cab9 · outbound

This paper cites acl-long.560/.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures acl-long.560/

Reference 560

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.368290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.368290Z digest=sha256:01997b181628af5ec01db5e83be6e3082163e4f69cca3a80b31b97f798ce8f6b

Observation 017e5cc6-e29c-4578-8d33-a18a77c18dd1 · outbound

This paper cites $” to ensure a unique structure. The final state is a pre-order traversal collecting edge labels, with child edges visited in lexicographical order and “$.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures $” to ensure a unique structure. The final state is a pre-order traversal collecting edge labels, with child edges visited in lexicographical order and “$

Reference 765

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:42:40.136462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:42:39.596354Z digest=sha256:01404109cc373ee8723763deae5bd230f268f6c6a8475b65c912ab119dadfe5b

Observation 22b4c08b-010c-4700-bf8e-3231496757eb · outbound

This paper cites Jain, N., Han, K., Gu, A., Li, W.-D., Yan, F., Zhang, T., Wang, S., Solar-Lezama, A., Sen, K., and Stoica, I.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Jain, N., Han, K., Gu, A., Li, W.-D., Yan, F., Zhang, T., Wang, S., Solar-Lezama, A., Sen, K., and Stoica, I

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:42:40.332011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:42:38.831254Z digest=sha256:f8bcfc6158010c48657eb87d1403b39580dccb0151f155b1ff5bd5a38e2e0392

Observation 42df52bc-c515-45b3-9f28-9356b6f805a8 · outbound

This paper cites Benchmark Data Contamination of Large Language Models: A Survey.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Benchmark Data Contamination of Large Language Models: A Survey

Reference 2022

Resolution
malformed identifier
no resolver link, observed 2026-08-07T12:42:39.495637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.495637Z digest=sha256:50886779c2acc6a1ff0b2d4470378d7712b2a3447949c42a269ce1050a7e54e9

Observation 762da522-2267-4dee-9448-65391b593145 · outbound

This paper cites doi: 10.1145/3564240.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures doi: 10.1145/3564240

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.260078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.260078Z digest=sha256:b842fa88662b873fdf87b9178e4abe8289f9eb93562f413cdfda1fa0b3be959f

Observation 46b88ea4-36b7-4e71-a9ad-d96715b6329a · outbound

This paper cites MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.012433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.012433Z digest=sha256:504917fda0a0927ddf1dbc38cbdd2752cd9035e50bbfebac84e0550acb597cff

Observation bc074de8-e6eb-408b-bfe5-ccef2fb27b58 · outbound

This paper cites Style Outweighs Substance: Failure Modes of LLM Judges in Alignment Benchmarking.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Style Outweighs Substance: Failure Modes of LLM Judges in Alignment Benchmarking

Reference 2025

Resolution
malformed identifier
no resolver link, observed 2026-08-07T12:42:38.753358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:38.753358Z digest=sha256:c5cd6e3a11da623e47712a829e1c132f86481d67c737f3349c06d3e5ba941af2

Pith citing papers

No inbound Pith citation observations are available.