Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:32:03.419764Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 6 inbound Pith citation observations for arXiv:2412.01333.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:32:03.419764Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:56:23.781576Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T03:38:57.833957Z
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e13550c2-d7d7-4b7b-911f-d01e9109d06b · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Summarizing source code using a neural attention model,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e07fe13-e465-4554-84b5-f2bad33e5b36 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? A neural model for generating natural language summaries of program subroutines,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 928dd8a9-7c7a-4be0-9553-650774b2da6d · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Large language models are few-shot summarizers: Multi-intent comment generation via in-context learning,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2791eb5d-5d3e-4c55-b31b-d6ed9e13f9ea · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Gpt-4 technical report,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bf09b46-b5e1-45af-9603-b472024bc83e · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Gemini: A Family of Highly Capable Multimodal Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5271e906-7a83-4617-a52f-a800452f994d · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? LLaMA: Open and Efficient Foundation Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fad18f81-5cb9-49ee-8f83-b744c0e6d6c3 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Bleu: a method for automatic evaluation of machine translation,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9c513bd-fada-4b2b-81ce-aa701e90a11c · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Rouge: A package for automatic evaluation of summaries,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dca472b4-82b1-4bc5-94d7-7afe777d0d5a · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34ee4429-ac5e-4887-887b-36d6a5629b3d · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? BERTScore: Evaluating Text Generation with BERT
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 483510d7-aebc-4846-b9ce-9a2b61aee303 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? On the evaluation of neural code summarization,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9b77032c-d91f-44a5-99a2-7b5f51adcf4c · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Reassessing automatic evaluation metrics for code summarization tasks,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac43f845-a8d3-4563-997b-b3718a9a75be · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Is ChatGPT a Good NLG Evaluator? A Preliminary Study
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9584fe44-95b9-4350-a50f-67e56050f521 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 035363bf-f34a-4f15-9bfe-707d3331831b · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Can Large Language Models Be an Alternative to Human Evaluations?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c6c2898-cd37-47a3-971f-33cbb43c2730 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? A Closer Look into Automatic Evaluation Using Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 165f2e09-5617-4951-b96f-947689a3593c · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af062056-6ed3-4a15-af6c-9b018670f92b · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Large Language Models Are State-of-the-Art Evaluators of Translation Quality
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d8c2549-bcb2-4420-8fd8-d78d3a0e977e · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Evaluating Large Language Models at Evaluating Instruction Following
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b52a792b-0557-4f99-8536-af24bb7b6196 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Ice-score: Instructing large language models to evaluate code,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e0a1c37-c890-4b52-95ee-5020050223ea · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Chatgpt,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e7d26a14-c503-436e-a99e-3a5e84e13a79 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? [Online]
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a56b8c3d-d6ea-4aed-b3af-c248cd382820 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? [Online]
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2382104c-c26d-4491-9956-ef2a9b8c270b · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Deep code comment generation,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b0bd9099-3c0d-4e2c-b09a-4b3f6d3719c3 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Retrieval-based neural source code summarization,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e99ece5-01bb-4805-b075-b70aa2f38490 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? A Transformer-based Approach for Source Code Summarization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a2c6c2d-e4f7-4cf7-bcb3-04eb15baa7c6 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Automatic Code Summarization via ChatGPT: How Far Are We?
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14c682dc-1c84-4b8f-99db-9e9b3a1b8895 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? An empirical study of smoothing techniques for language modeling,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e1ede7c-79e7-49a3-9d80-c5b425441890 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Retrieve and refine: exemplar- based neural comment generation,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fcf31da2-c731-49b8-907d-7ff8a7e96804 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Project-level encoding for neural source code summarization of subroutines,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 596247f9-e949-4d1b-8934-b6c2cf99f0d9 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Improving code summarization with block-wise abstract syntax tree splitting,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 39323a8a-9b66-42b9-9826-e7149f3afc6f · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Api2com: On the improvement of automatically generated code comments using api documentations,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5d97a36b-2e55-46ba-88fd-b783ed0f6cd5 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3835d723-d7fd-4849-b663-91fac5d2e354 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Chain-of-thought prompting elicits reasoning in large language models,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e27429a5-6910-42c0-bc4d-a1b999f7217b · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Generative agents: Interactive simulacra of human behavior,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1255f785-8cff-488e-b14c-118bf72ed027 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Multi-Agent Consensus Seeking via Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ad4f033-cd94-4a46-a002-08c98a7a4db0 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Summarizing source code with transferred api knowledge,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a251f891-ac49-4969-81ed-68eef0843c67 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Summeval: Re-evaluating summarization evaluation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 242d043c-459e-4a45-9221-023517833854 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? The treatment of ties in ranking problems,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 97189272-8706-4ffe-b1a7-2296e3c32757 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Standard probability and statistics tables and formulae,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6ccc1155-2c2c-435a-9e6d-383e9291aa9a · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Software documentation: how much is enough?
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d83cfa2d-63cc-4301-846a-d13cee2f0c6f · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? The relevance of software documenta- tion, tools and technologies: a survey,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6a7268c7-3224-4de3-bf59-274f3c943eb3 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Documenting software systems with views,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6a13a4b6-cb70-4382-8cc2-bab14cb0c6ec · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Reinforcement-learning-guided source code summarization using hierarchical attention,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5c414b7d-351d-47ae-8c67-168d83bec955 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Modeling hierarchical syntax structure with triplet position for source code summarization,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0899d916-3e40-4623-8508-78afb46fe682 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Improving automatic source code summarization via deep reinforcement learning,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 55ed83f5-0be1-477c-bb60-4d24247c8147 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Deep learning for code intelligence: Survey, benchmark and toolkit,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 25fedac6-7b5f-489b-bbb3-416a3861b673 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Deep learning for code generation: A survey,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 129d47aa-c4cf-4eab-b7af-c1fdac64973e · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Code Summarization with Structure-induced Transformer
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 519c9f0c-4e48-4f07-9762-c7d5d9ae4bd9 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? code2seq: Generating Sequences from Structured Representations of Code
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af9034e2-1d4a-49c3-8a0c-b4337a41265d · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Code to comment
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5ea8343-2320-4df4-9fde-08f82402630e · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? CodeBERT: A Pre-Trained Model for Programming and Natural Languages
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b802330c-1d96-4284-8feb-47a0654719d1 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 908919ee-ec26-4862-a5ef-61ac948d13ad · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Automatic code documentation generation using gpt-3,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 59d5ab25-a658-47f0-970e-320985656d5f · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Automatic semantic augmentation of language model prompts (for code summarization),
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1190defe-72f6-48eb-a264-9120e79caed6 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? On-the-fly adapting code summarization on trainable cost-effective language models,
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f9d36c9-3c8a-4e7c-a72d-119365f9c683 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Distilled gpt for source code summarization,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f229d004-d995-46ea-b242-f0c9bc8efe11 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Source code summarization in the era of large language models,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 918d554c-29ff-42b0-a47f-eb040aafcb44 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? A human study of comprehension and code summariza- tion,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fcb3384d-e882-4cc3-9be6-cc3a2a7c7699 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? GPTScore: Evaluate as You Desire
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04e8875f-7143-4f7b-a991-1ad53d92a41a · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Mllm-as-a-judge: Assessing multimodal llm-as- a-judge with vision-language benchmark,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 78ad570d-338a-4aef-8e62-c34cbd16e3d7 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Large Language Models are Diverse Role-Players for Summarization Evaluation
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab4ed19b-f959-4989-8ddc-62fd7bb645e8 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? How Reliable Are Automatic Evaluation Methods for Instruction-Tuned LLMs?
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f94e10d2-e19b-4132-ba34-ec66c88881c6 · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Who validates the validators? aligning llm-assisted evaluation of llm outputs with human preferences,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3ad261fa-84c8-4368-8b52-5d20790a899a · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Naturalcc: an open-source toolkit for code intelligence,
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation aeb14d77-2f83-45cd-a000-5c29818e25cc · outbound
Can Large Language Models Serve as Evaluators for Code Summarization? Available: https://doi.org/10.1145/3654777.3676450
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4339c255-4ae4-4a99-8527-11901eb565f7 · inbound
Can Large Language Models Understand Intermediate Representations in Compilers? Can Large Language Models Serve as Evaluators for Code Summarization?
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e6845cf-8bd7-4310-b6a8-9d2d244f04ae · inbound
Do Automatic Comment Generation Techniques Fall Short? Exploring the Influence of Method Dependencies on Code Understanding Can Large Language Models Serve as Evaluators for Code Summarization?
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bd6ec4a-2d93-4b15-8bff-a2804792981c · inbound
Issue Retrieval and Verification Enhanced Supplementary Code Comment Generation Can Large Language Models Serve as Evaluators for Code Summarization?
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6813f7d-0b7f-4e78-8e46-f4d83a39b5ce · inbound
evalSmarT: An LLM-Based Framework for Evaluating Smart Contract Generated Comments Can Large Language Models Serve as Evaluators for Code Summarization?
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbca0308-c386-41ea-b3d0-3003a1157511 · inbound
Knowledge-Graph-Driven Data Synthesis for Low-Resource Software Development: A HarmonyOS Case Study Can Large Language Models Serve as Evaluators for Code Summarization?
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6c19b454-d264-47db-a20e-6d6c0a437ead · inbound
Using Mutation-Analysis to Examine an LLM's Ability to Summarize Code Can Large Language Models Serve as Evaluators for Code Summarization?
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.