Pith. sign in

Paper Citation Record · LEDGER

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures

As of 7 August 2026, this Paper Citation Record lists 9 of 9 outbound references and 0 inbound Pith citation observations for arXiv:2505.24069.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24069 v4

Coverage vector

measured 9 of 9 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:42:39.596354Z

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

9 of 9 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved5
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c931fb64-0f01-4a99-a450-5eff50ef3507 · outbound

This paper cites Mixtral of Experts.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Mixtral of Experts

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:38.862158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:38.862158Z digest=sha256:d9db9ea8bf9a5c947de0ec334a398a7aae08c12ca108be3db821a8f609bac583

Observation f9b46fef-c915-48e2-8cf3-c2d7b4a95500 · outbound

This paper cites Code Simulation Challenges for Large Language Models.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Code Simulation Challenges for Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.128432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.128432Z digest=sha256:115eb92e50318061f2efca6c28fb8635688280709c1d9d27817d715bc545e69d

Observation 520f27ae-e1ca-41b4-9077-2fdfa797cab9 · outbound

This paper cites acl-long.560/.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures acl-long.560/

Reference 560

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.368290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.368290Z digest=sha256:42277b95b184a965ccb8f639a6968acb3c6d985f4960e0435cb0b1d6aa1962e6

Observation 017e5cc6-e29c-4578-8d33-a18a77c18dd1 · outbound

This paper cites $” to ensure a unique structure. The final state is a pre-order traversal collecting edge labels, with child edges visited in lexicographical order and “$.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures $” to ensure a unique structure. The final state is a pre-order traversal collecting edge labels, with child edges visited in lexicographical order and “$

Reference 765

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:42:40.136462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:42:39.596354Z digest=sha256:b5d1430d9c644a15fd7c3e6a9a9f293fcd06a01e370acf690aca18623710d2b5

Observation 22b4c08b-010c-4700-bf8e-3231496757eb · outbound

This paper cites Jain, N., Han, K., Gu, A., Li, W.-D., Yan, F., Zhang, T., Wang, S., Solar-Lezama, A., Sen, K., and Stoica, I.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Jain, N., Han, K., Gu, A., Li, W.-D., Yan, F., Zhang, T., Wang, S., Solar-Lezama, A., Sen, K., and Stoica, I

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:42:40.332011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:42:38.831254Z digest=sha256:24cd4cbd3ae1232ae8c720e4122835f4316495e41467338d3ea6479de9dfd69e

Observation 42df52bc-c515-45b3-9f28-9356b6f805a8 · outbound

This paper cites Benchmark Data Contamination of Large Language Models: A Survey.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Benchmark Data Contamination of Large Language Models: A Survey

Reference 2022

Resolution
malformed identifier
no resolver link, observed 2026-08-07T12:42:39.495637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.495637Z digest=sha256:12c21d70787ee83bfbf5a07c9ebe8c562e501dc267d1b9c727eb7caf1c2b26f5

Observation 762da522-2267-4dee-9448-65391b593145 · outbound

This paper cites doi: 10.1145/3564240.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures doi: 10.1145/3564240

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.260078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.260078Z digest=sha256:a8dab60ea0bb391ca5b3e030441776768ed81ae9bfbb643c9899b8965db8c6ee

Observation 46b88ea4-36b7-4e71-a9ad-d96715b6329a · outbound

This paper cites MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:39.012433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:39.012433Z digest=sha256:a48f9c7a1a1b103dc19d91518a668ab1b65f432337f12409165303c8b1a0592b

Observation bc074de8-e6eb-408b-bfe5-ccef2fb27b58 · outbound

This paper cites Style Outweighs Substance: Failure Modes of LLM Judges in Alignment Benchmarking.

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures Style Outweighs Substance: Failure Modes of LLM Judges in Alignment Benchmarking

Reference 2025

Resolution
malformed identifier
no resolver link, observed 2026-08-07T12:42:38.753358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:38.753358Z digest=sha256:2ba4141e53953894645bfd10feb573dbfa21214efaeb2c6f5717f4b7f19c2dd7

Pith citing papers

No inbound Pith citation observations are available.