Pith. sign in

Paper Citation Record · LEDGER

A Code Comprehension Benchmark for Large Language Models for Code

As of 21 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.10641.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10641 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:36:56.819266Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5c981be3-ad04-4c41-8589-dcc141133f45 · outbound

This paper cites Supporting code comprehension via annotations: Right information at the right time and place.

A Code Comprehension Benchmark for Large Language Models for Code Supporting code comprehension via annotations: Right information at the right time and place

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:57.014846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.764675Z digest=sha256:35ec7430db30a6cc2f95b408c6583394901d984ce4de63d9d68a779ff8446ee7

Observation 8e45f332-405d-43be-9d2c-e38a0a8f301e · outbound

This paper cites SweLL Benchmark: Semantic Code Evaluation Beyond Generation.

A Code Comprehension Benchmark for Large Language Models for Code SweLL Benchmark: Semantic Code Evaluation Beyond Generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:57.005501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.768228Z digest=sha256:41d8b238b14f6ac98b9e0010c86931d8793d81341653106d942a022bf01a0749

Observation 0aa67325-f0c1-4280-b62b-569add226dc4 · outbound

This paper cites an unresolved cited work.

A Code Comprehension Benchmark for Large Language Models for Code Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:36:56.996609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.771338Z digest=sha256:23fe0a014688f72808e3f9f72855bf735024b5ce35190a3ff51cbeae9831d04e

Observation c4414c5d-878d-4b0f-a861-62e53e47095d · outbound

This paper cites ROUGE Score.

A Code Comprehension Benchmark for Large Language Models for Code ROUGE Score

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.988339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.773941Z digest=sha256:05f9203c69e2ad119e0692b4696266d48bb5d06090a5f8b1ab817bd10480544d

Observation b7548818-92b1-4dbb-9655-46c7b2513c9b · outbound

This paper cites Program Synthesis with Large Language Models.

A Code Comprehension Benchmark for Large Language Models for Code Program Synthesis with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.776968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.776968Z digest=sha256:39bbf042be3a4205039d9a2b500ede13f593c1e396a86e7ccc657501869622f8

Observation 0a38eba4-7529-4cee-b035-60df773e6379 · outbound

This paper cites Neural code comprehension: A learnable representation of code semantics.

A Code Comprehension Benchmark for Large Language Models for Code Neural code comprehension: A learnable representation of code semantics

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.979468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.780085Z digest=sha256:f72611e6aec398e4c497b5091adda5f227a1556788dd934395e431f1418d86fd

Observation c207355b-463c-4ee0-bb11-f45f6d7f0044 · outbound

This paper cites Fold2vec: Towards a statement-based representation of code for code comprehension.

A Code Comprehension Benchmark for Large Language Models for Code Fold2vec: Towards a statement-based representation of code for code comprehension

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.970509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.782978Z digest=sha256:42ff511dd1a18feab830e0320ffb1cfbbcd3269a06e84dbdc3a5f74b27815d54

Observation a85426fe-e4d1-4568-89ce-27c84982e9ca · outbound

This paper cites Evaluating Large Language Models Trained on Code.

A Code Comprehension Benchmark for Large Language Models for Code Evaluating Large Language Models Trained on Code

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.785601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.785601Z digest=sha256:62a7efd473592ca1d23854101df9c173015a523879940bf0b40af8575ff2c0f7

Observation a8280d3c-b8e9-49c7-8d26-6b19e83fff87 · outbound

This paper cites Z3: An efficient SMT solver.

A Code Comprehension Benchmark for Large Language Models for Code Z3: An efficient SMT solver

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.961935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.788351Z digest=sha256:4bbf68ec744c7104b8a0fafbdfb9c3f2b3d57424b8efa2d963c733322b8389f9

Observation f12f33d5-eb39-4dec-84df-e807f105a3e8 · outbound

This paper cites A comprehensive review on software comprehension models.

A Code Comprehension Benchmark for Large Language Models for Code A comprehensive review on software comprehension models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.953345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.790740Z digest=sha256:14e0c34d1c6b8711e902f3a669db0f2fa4ffed23b558dadd75869fedf366700f

Observation 34b5492c-9497-4625-a336-601b3ae25931 · outbound

This paper cites Neural semantic parsing.

A Code Comprehension Benchmark for Large Language Models for Code Neural semantic parsing

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.943832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.793407Z digest=sha256:2da7cd61145345838cfd5a9249b8808cc588467c6692a54f524920378d36a4d0

Observation e3049464-4672-46ed-9886-c960a6469319 · outbound

This paper cites A formalism for dependency grammar based on tree adjoining grammar.

A Code Comprehension Benchmark for Large Language Models for Code A formalism for dependency grammar based on tree adjoining grammar

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.935243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.795789Z digest=sha256:e9fe20c24e339f168499c64fe5cb4962bab417e6df39d50c4ac7059548513ccf

Observation 89320922-a079-4270-bb58-75ca13ed7567 · outbound

This paper cites StarCoder: may the source be with you!.

A Code Comprehension Benchmark for Large Language Models for Code StarCoder: may the source be with you!

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.798322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.798322Z digest=sha256:1490b21c4dd30c49780a6e8accad019c3f23b54c8b39b8aa3d68f821708f5781

Observation bdcb7977-b4d0-4239-92fe-37d111378efb · outbound

This paper cites Competition-level code generation with alphacode.

A Code Comprehension Benchmark for Large Language Models for Code Competition-level code generation with alphacode

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.925514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.801032Z digest=sha256:7646fab6fdb377d162929d21f8152baf57eeae20b61ecebe7cbd8c468ba0fc44

Observation 4f26b513-5ca3-400a-8081-f2a4963df00c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

A Code Comprehension Benchmark for Large Language Models for Code LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.803347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.803347Z digest=sha256:608311bdbbf664294bd248e31279459d83735f1991f401ad8757650ed435e6ba

Observation b2d67574-3e18-411d-989d-3055e153dca4 · outbound

This paper cites Studying the usage of text-to-text transfer transformer to support code- related tasks.

A Code Comprehension Benchmark for Large Language Models for Code Studying the usage of text-to-text transfer transformer to support code- related tasks

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.916649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.805909Z digest=sha256:9e00888f316571370e6cdbef572903fe6cdb40572312165ecfe8c76c1a1e1c26

Observation 6b5864e9-359f-42b4-a724-efa162557670 · outbound

This paper cites Barriers for students during code change com- prehension.

A Code Comprehension Benchmark for Large Language Models for Code Barriers for students during code change com- prehension

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.908849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.808341Z digest=sha256:5e147f76964c8ee99bc1eeccb18b05adf8fb3e03838e49f8b74650c30c32dfad

Observation 29004c6c-2bba-4442-87f6-29c8de30abfb · outbound

This paper cites CodeBLEU: a Method for Automatic Evaluation of Code Synthesis.

A Code Comprehension Benchmark for Large Language Models for Code CodeBLEU: a Method for Automatic Evaluation of Code Synthesis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T17:36:56.810676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:36:56.810676Z digest=sha256:b5f456564bdbb6cae2adbc47992abb5480b4c87fdca8169cd2b3e92569f27080

Observation c3e53d8f-273c-4694-bfdd-fa790b0a3ef9 · outbound

This paper cites An empirical approach to understand the role of emotions in code comprehension.

A Code Comprehension Benchmark for Large Language Models for Code An empirical approach to understand the role of emotions in code comprehension

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.901100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.814059Z digest=sha256:c3b4d2cb0dbfe0814f54eb2c007cf8201a7184d98d756892510d395a64889dd7

Observation 569b90ab-3ad2-4e23-a91c-9ea4909ebe5d · outbound

This paper cites An empirical study on learning bug-fixing patches in the wild via neural machine translation.

A Code Comprehension Benchmark for Large Language Models for Code An empirical study on learning bug-fixing patches in the wild via neural machine translation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.893229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.816328Z digest=sha256:b492c5dab7d02d5c6072974d12bc1aef614dc639dafd952898a9419100383971

Observation 5372a766-afd8-4d7e-add6-f50de1b1bb4e · outbound

This paper cites https: //github.com/tongye98/Awesome-Code-Benchmark.

A Code Comprehension Benchmark for Large Language Models for Code https: //github.com/tongye98/Awesome-Code-Benchmark

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:36:56.885128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T17:36:56.819266Z digest=sha256:48717be2d0a0e6a15509614875e574dcd4834a691c82b916b5054ca8b50397c6

Pith citing papers

No inbound Pith citation observations are available.