Pith. sign in

Paper Citation Record · LEDGER

A Code Comprehension Benchmark for Large Language Models for Code

As of 8 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.10641.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10641 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:36:56.819266Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5c981be3-ad04-4c41-8589-dcc141133f45 · outbound

This paper cites Supporting code comprehension via annotations: Right information at the right time and place.

A Code Comprehension Benchmark for Large Language Models for Code Supporting code comprehension via annotations: Right information at the right time and place

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:57.014846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.764675Z digest=sha256:eddc48113bc0edeee5da9740b9414fbcd94db40d67cd06bb3128c649211a37e9

Observation 8e45f332-405d-43be-9d2c-e38a0a8f301e · outbound

This paper cites SweLL Benchmark: Semantic Code Evaluation Beyond Generation.

A Code Comprehension Benchmark for Large Language Models for Code SweLL Benchmark: Semantic Code Evaluation Beyond Generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:57.005501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.768228Z digest=sha256:cddca7e32b038fb31a0f662c83585ab4dcee78c963bd04a273241c1aeaf4f472

Observation 0aa67325-f0c1-4280-b62b-569add226dc4 · outbound

This paper cites an unresolved cited work.

A Code Comprehension Benchmark for Large Language Models for Code Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:36:56.996609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.771338Z digest=sha256:8bc8db764e30ea3bce06d60a74e2e562fa42f1cf29aa84bc7687a8fb80ef5b4c

Observation c4414c5d-878d-4b0f-a861-62e53e47095d · outbound

This paper cites ROUGE Score.

A Code Comprehension Benchmark for Large Language Models for Code ROUGE Score

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.988339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.773941Z digest=sha256:f6559776ed335a2f5e77cc56dfc0132511f20b28cb99099dbe5b249a31e32986

Observation b7548818-92b1-4dbb-9655-46c7b2513c9b · outbound

This paper cites Program Synthesis with Large Language Models.

A Code Comprehension Benchmark for Large Language Models for Code Program Synthesis with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.776968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.776968Z digest=sha256:fbe12a7569ac8a3043ba4b904bc023a66b38270e4116a838eacd5d2e4ecd448b

Observation 0a38eba4-7529-4cee-b035-60df773e6379 · outbound

This paper cites Neural code comprehension: A learnable representation of code semantics.

A Code Comprehension Benchmark for Large Language Models for Code Neural code comprehension: A learnable representation of code semantics

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.979468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.780085Z digest=sha256:1f87b06379924e1c736e8d3a01cb1650c022a24324aa5cbc7ba2695547215264

Observation c207355b-463c-4ee0-bb11-f45f6d7f0044 · outbound

This paper cites Fold2vec: Towards a statement-based representation of code for code comprehension.

A Code Comprehension Benchmark for Large Language Models for Code Fold2vec: Towards a statement-based representation of code for code comprehension

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.970509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.782978Z digest=sha256:6258ac130b0307bfedb13dcabe49b5247e8c061d42a8789bb9c649100e170f9c

Observation a85426fe-e4d1-4568-89ce-27c84982e9ca · outbound

This paper cites Evaluating Large Language Models Trained on Code.

A Code Comprehension Benchmark for Large Language Models for Code Evaluating Large Language Models Trained on Code

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.785601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.785601Z digest=sha256:9153a8c29e9e7cc56524da9ebe92cea0fe40686b7c583b0c46f21155cfa84c7d

Observation a8280d3c-b8e9-49c7-8d26-6b19e83fff87 · outbound

This paper cites Z3: An efficient SMT solver.

A Code Comprehension Benchmark for Large Language Models for Code Z3: An efficient SMT solver

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.961935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.788351Z digest=sha256:9af5543410cf1639f5ec98846f2e6510cd7e3951497f3d870279b1e0d33b87d0

Observation f12f33d5-eb39-4dec-84df-e807f105a3e8 · outbound

This paper cites A comprehensive review on software comprehension models.

A Code Comprehension Benchmark for Large Language Models for Code A comprehensive review on software comprehension models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.953345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.790740Z digest=sha256:2a0d2b984aa254eb794283e298fbeb938c4514d6528b551509ac7116f0c0aef2

Observation 34b5492c-9497-4625-a336-601b3ae25931 · outbound

This paper cites Neural semantic parsing.

A Code Comprehension Benchmark for Large Language Models for Code Neural semantic parsing

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.943832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.793407Z digest=sha256:08cc22375730c53dbbd9e04fbde09ec848c93e1953b6a615e52c3ddc0fcb6696

Observation e3049464-4672-46ed-9886-c960a6469319 · outbound

This paper cites A formalism for dependency grammar based on tree adjoining grammar.

A Code Comprehension Benchmark for Large Language Models for Code A formalism for dependency grammar based on tree adjoining grammar

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.935243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.795789Z digest=sha256:26c5f0e58285d827a795494bd8cb2802f6941475e6a36253e19eaa16eb30436f

Observation 89320922-a079-4270-bb58-75ca13ed7567 · outbound

This paper cites StarCoder: may the source be with you!.

A Code Comprehension Benchmark for Large Language Models for Code StarCoder: may the source be with you!

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.798322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.798322Z digest=sha256:1490b21c4dd30c49780a6e8accad019c3f23b54c8b39b8aa3d68f821708f5781

Observation bdcb7977-b4d0-4239-92fe-37d111378efb · outbound

This paper cites Competition-level code generation with alphacode.

A Code Comprehension Benchmark for Large Language Models for Code Competition-level code generation with alphacode

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.925514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.801032Z digest=sha256:f9440f341332893997ff69399a4dc359b01aefaf07ec48fc22fc01484727206c

Observation 4f26b513-5ca3-400a-8081-f2a4963df00c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

A Code Comprehension Benchmark for Large Language Models for Code LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.803347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.803347Z digest=sha256:e6ead5fa3071f64c2a493838d7eaf9d41b5762a5a26630cb7b52f27ef4fd20f5

Observation b2d67574-3e18-411d-989d-3055e153dca4 · outbound

This paper cites Studying the usage of text-to-text transfer transformer to support code- related tasks.

A Code Comprehension Benchmark for Large Language Models for Code Studying the usage of text-to-text transfer transformer to support code- related tasks

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.916649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.805909Z digest=sha256:ad1fcb7e8acbced5654894b21fcce05ed64c2f3f1ea9e89ce919e2588e5b1c77

Observation 6b5864e9-359f-42b4-a724-efa162557670 · outbound

This paper cites Barriers for students during code change com- prehension.

A Code Comprehension Benchmark for Large Language Models for Code Barriers for students during code change com- prehension

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.908849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.808341Z digest=sha256:2c47804b32943798a0b732fc0171600111e00f3542bbd3cc481f4dcc03f3b493

Observation 29004c6c-2bba-4442-87f6-29c8de30abfb · outbound

This paper cites CodeBLEU: a Method for Automatic Evaluation of Code Synthesis.

A Code Comprehension Benchmark for Large Language Models for Code CodeBLEU: a Method for Automatic Evaluation of Code Synthesis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.810676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.810676Z digest=sha256:81dae48e277b689d7bff40a4992c4a20d09010a3509362228f29184564bedb80

Observation c3e53d8f-273c-4694-bfdd-fa790b0a3ef9 · outbound

This paper cites An empirical approach to understand the role of emotions in code comprehension.

A Code Comprehension Benchmark for Large Language Models for Code An empirical approach to understand the role of emotions in code comprehension

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.901100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.814059Z digest=sha256:71bb2b8b351c698b47054ce850e2080a1d633c49b8cf877b6a1ad5c05f3e6479

Observation 569b90ab-3ad2-4e23-a91c-9ea4909ebe5d · outbound

This paper cites An empirical study on learning bug-fixing patches in the wild via neural machine translation.

A Code Comprehension Benchmark for Large Language Models for Code An empirical study on learning bug-fixing patches in the wild via neural machine translation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.893229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.816328Z digest=sha256:1bb2ca509e7daf04f89295629eb9cde2c18ace737b358eca767876fd5ee6fa30

Observation 5372a766-afd8-4d7e-add6-f50de1b1bb4e · outbound

This paper cites https: //github.com/tongye98/Awesome-Code-Benchmark.

A Code Comprehension Benchmark for Large Language Models for Code https: //github.com/tongye98/Awesome-Code-Benchmark

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.885128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:36:56.819266Z digest=sha256:1b9aadb0e0a80108219fc910aea43be7a6583f1f2f158c297fb27409f7e8dd86

Pith citing papers

No inbound Pith citation observations are available.