Pith. sign in

Paper Citation Record · LEDGER

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension

As of 21 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2412.00314.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00314 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:35:51.340765Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:24:28.140998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T16:24:28.418595Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 752027d0-541b-4e23-9dde-d15fe79095bc · outbound

This paper cites Who judges the judge: An empirical study on online judge tests,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Who judges the judge: An empirical study on online judge tests,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:52.039866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.117340Z digest=sha256:f570ec62a6df3eedcfa60c4988a1d2bf32b8bfe0c8b98b92e5cb41ea2cfeb3db

Observation f02cab78-82c1-40ef-93de-817ed722b6ba · outbound

This paper cites Automatic source code evaluation test develop- ment in programming education using grey-box methods,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Automatic source code evaluation test develop- ment in programming education using grey-box methods,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:52.022557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.123217Z digest=sha256:54e1d39aa539a9078239523d8201cb27195332aea0f63ca15d6d9dbc35ad1f2e

Observation e6d004e8-46e5-4149-a0f6-ebf5bbbd1cc2 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.128443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.128443Z digest=sha256:834a4332b3c848bba5087bcafcc3797ce5c8074dcac786a9ded8baaac2ccbde6

Observation f6f2a5a7-d864-428e-b609-d78100806a24 · outbound

This paper cites Coderl: Mastering code generation through pretrained models and deep reinforcement learning,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Coderl: Mastering code generation through pretrained models and deep reinforcement learning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.133846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.133846Z digest=sha256:01b52dc95fb956b348c48fe256b57c3716b631da5a593917d555081a11b6283d

Observation deb65a93-2211-412d-927b-f870e39d239f · outbound

This paper cites Large language models meet nl2code: A survey,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Large language models meet nl2code: A survey,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.993817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.139127Z digest=sha256:4b17cdeec7a4807a53aa6d1e797063366084a465378f5bc47b9bc87f163af228

Observation 1cd7d9a5-1049-42db-9821-3db990e28abc · outbound

This paper cites CodeScore: Evaluating Code Generation by Learning Code Execution.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeScore: Evaluating Code Generation by Learning Code Execution

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.144229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.144229Z digest=sha256:d04f949b704509a930ce49b314c6b7bb0eac7a2b8817a2d5b45b7d7242991e0c

Observation 3992f1cc-c07c-4b5e-aa1b-3b96311eb678 · outbound

This paper cites Spoc: Search-based pseudocode to code,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Spoc: Search-based pseudocode to code,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.150080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.150080Z digest=sha256:edeb190f051d2ab70f4b9cc675604b0434b3001bf196868751d428917f210e16

Observation c0bd294d-bee7-4eb8-a74d-6da231ae2dd3 · outbound

This paper cites Measuring Coding Challenge Competence With APPS.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Measuring Coding Challenge Competence With APPS

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.154868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.154868Z digest=sha256:f35b1fba719b09056597295f5771d74ebf1e1a5b59d53ba01904ac3fa2c2850a

Observation 1a5342d1-b826-470c-ab36-4e52373ada91 · outbound

This paper cites Out of the bleu: how should we assess quality of the code generation models?.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Out of the bleu: how should we assess quality of the code generation models?

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.966553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.160198Z digest=sha256:1c8d8d4c92db47d6fad940cd5cd8a6cc8316f3f4386a6bb631799f588091bbed

Observation 1287646d-da8b-41d5-8c8c-112888733f82 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Bleu: a method for automatic evaluation of machine translation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.164815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.164815Z digest=sha256:ea8f7b7c9b4cd21f0424354aa5c94905cd48f69127b6ce8844fe58914e83a662

Observation c60260b5-3f25-40ea-a253-c423b5844166 · outbound

This paper cites A package for automatic evaluation of summaries,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension A package for automatic evaluation of summaries,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.938577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.169661Z digest=sha256:5cf1b1e4e6169edd4739cb3ecb0f11a739d91068c3c5c923f662989e6ba9b91b

Observation cdd5a12a-ef2c-468e-ac8c-754a3cc599ff · outbound

This paper cites Does bleu score work for code migration?.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Does bleu score work for code migration?

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.921968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.174483Z digest=sha256:b6dab0e95e149a1d9d06948f207e4dd8a07d6e62998f9ee82b96f8423121ce11

Observation f5dc2978-eab1-4b94-8c2f-b83e9be49e99 · outbound

This paper cites CodeBERTScore: Evaluating Code Generation with Pretrained Models of Code.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeBERTScore: Evaluating Code Generation with Pretrained Models of Code

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.179925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.179925Z digest=sha256:0b08fb6346f4e6f7eddb0ae223080d975ddab8023817c49740428b4c73db813a

Observation 068580ea-0365-4a79-8662-dc320204b5af · outbound

This paper cites CodeBLEU: a Method for Automatic Evaluation of Code Synthesis.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeBLEU: a Method for Automatic Evaluation of Code Synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.185496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.185496Z digest=sha256:ca6a42111be9b422bc758c408f9be75c5868b75b2646dc3f28f9ed10d5074600

Observation b00d2833-ff52-4abf-b2af-97f2860c9cfc · outbound

This paper cites chrf: character n-gram f-score for automatic mt evalu- ation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension chrf: character n-gram f-score for automatic mt evalu- ation,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.190751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.190751Z digest=sha256:abd7c53aa3e2c6d33f4bd0648aa63d3065faae6cf0934b4d069e49bccdd5755a

Observation 526d3951-d78a-4f68-a0a5-34915fb4f34a · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.195631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.195631Z digest=sha256:3aadbc628f97b4fbc6dd20e938e628d9edb1ee695d3c661298b805bf46cfacbf

Observation a931a814-1e50-4419-8454-83330b2aedd9 · outbound

This paper cites ICE-Score: Instructing Large Language Models to Evaluate Code.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension ICE-Score: Instructing Large Language Models to Evaluate Code

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.200654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.200654Z digest=sha256:b72bbf593c3909228dbafebc2f6d83fc796baca0cbe2d3428afa4e17ed475665

Observation fc582aaf-96ce-489c-bc10-abc4196325ca · outbound

This paper cites Learning to mine aligned code and natural language pairs from stack overflow,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Learning to mine aligned code and natural language pairs from stack overflow,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.206144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.206144Z digest=sha256:95fae85da01b8da6bb196e264e2d4c64cf1f799e586cfa8313097ac6128d55a5

Observation dc7885f2-297a-4488-b3a4-02136c718eaa · outbound

This paper cites Latent Predictor Networks for Code Generation.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Latent Predictor Networks for Code Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.211234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.211234Z digest=sha256:71e40f63fd65fe89b9aad6523f5588e4b0b1ac90febcc90852af0eaa3415d91f

Observation 7b7dda7c-3b6a-4183-a7ec-cd239a5faef4 · outbound

This paper cites Towards a Unified Multi-Dimensional Evaluator for Text Generation.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.217643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.217643Z digest=sha256:f0b9df84925eae512b1123cb4044edc45813903f68396006f21cf4903202d067

Observation 9b1269a7-94a0-469d-ba27-29e26808068e · outbound

This paper cites Survey of hallucination in natural language generation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Survey of hallucination in natural language generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.223383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.223383Z digest=sha256:3072b2e30b39fdc71436b49b03ee85c979669f29363f5c4759c29f397d83a5cf

Observation 911bbbca-0825-4cfd-aebe-06eb822312c4 · outbound

This paper cites Faithful Reasoning Using Large Language Models.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Faithful Reasoning Using Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.228537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.228537Z digest=sha256:cf2eb18943c908c03c9d2c6a8420f051c560b3dda8bc353979651c76ff6f1c7f

Observation 38d12603-7d58-4eef-b6d8-a1b436fc1f86 · outbound

This paper cites CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.233779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.233779Z digest=sha256:57acb85a54bb0b04c7738ae0fc67c51031792f6704fa3636a3b2d80153c7e977

Observation 15094510-9d6d-4464-adee-e218b43d67e8 · outbound

This paper cites Increasing employability of indian engineering graduates through experiential learning programs and competitive programming: Case study,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Increasing employability of indian engineering graduates through experiential learning programs and competitive programming: Case study,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.871489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.239284Z digest=sha256:02c3e022ddf0cd853e3ac643c21f932d636b9832e2525bc07668397e58361de0

Observation d119ebda-43a2-480f-8312-099250ea6d70 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Evaluating Large Language Models Trained on Code

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.244040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.244040Z digest=sha256:58b62677e66792763406cd745812d6f70b19eea8965c694483aa58376fd6d8cc

Observation a7846e2d-37fc-4255-bb1b-7869fa9d4059 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Chain-of-thought prompting elicits reasoning in large language models,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.249521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.249521Z digest=sha256:cd17060c9f920607cf6c2289206ae2c1f39efe2b7797680376d12f584444934a

Observation faa84c3d-c7b7-4c30-b14c-c3d58d8170ad · outbound

This paper cites Fine-grained code clone detection with block-based splitting of abstract syntax tree,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Fine-grained code clone detection with block-based splitting of abstract syntax tree,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.843479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.254513Z digest=sha256:a34f5f110e8a50e7a64b9cd580c1e5d8eba26eebb031e21e88ec4499bb06812a

Observation 4cdc6928-b329-4be6-a59f-0a6e3d58de07 · outbound

This paper cites Cocoast: Representing source code via hierarchical splitting and re- construction of abstract syntax trees,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Cocoast: Representing source code via hierarchical splitting and re- construction of abstract syntax trees,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.827289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.259535Z digest=sha256:02c9be1437211885a249ab9a707af5a3b4cf5673892105c53e3d52037c061574

Observation d0f395a4-6c88-4b23-8d2e-8e074dec6c5f · outbound

This paper cites Blocsum: Block scope-based source code summarization via shared block representation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Blocsum: Block scope-based source code summarization via shared block representation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.809874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.264836Z digest=sha256:a2565f8537c6bec6105ec23aed681e3db3dbbbb867ec9949b10c68c65cc59b3c

Observation 674e58ec-730e-4b40-b7df-0ba2037027af · outbound

This paper cites Openai gpt-3.5 turbo,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Openai gpt-3.5 turbo,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.791428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.270072Z digest=sha256:e9422e4e5958c99a67ce3c2ca2e4e050eba57428f1e6f4d9f8c92f4bc179af3f

Observation 2aeb5347-8c25-45fc-bcd5-f766c271daf2 · outbound

This paper cites Exploring security vulnerabilities in competitive programming: An empirical study,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Exploring security vulnerabilities in competitive programming: An empirical study,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.774773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.275147Z digest=sha256:4287571898dd60a9ca5973e13c56e78d9f3f07a26da283f6bf0bbafb59f2237e

Observation 3ae686d0-9386-4566-837b-0b9fd98375a2 · outbound

This paper cites Multipl-e: A scalable and polyglot approach to benchmarking neural code generation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Multipl-e: A scalable and polyglot approach to benchmarking neural code generation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.756582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.280106Z digest=sha256:50bbb10cf9b50f3ee283f7656102ad90d152b9aedb4620349c311fead462cb46

Observation 69553955-b7c8-4108-a109-f3161c4fe0c0 · outbound

This paper cites Gpt-4: Language models at scale,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Gpt-4: Language models at scale,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.739906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.285090Z digest=sha256:a7eaf82e60d8ecbb125bbb7a6927f819c74fad3a79e9ce34bed14848f3057f45

Observation 7d8c3c11-2f7e-4ff2-a96a-7a5e878a1e38 · outbound

This paper cites Openai gpt-4 turbo,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Openai gpt-4 turbo,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.722040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.289829Z digest=sha256:982f11391c8284f32fded4a9f05f99c389cb6e07ef172efa614831f5b1ac6e29

Observation 7c7178db-2b09-486a-aa18-98ba39063488 · outbound

This paper cites Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.294747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.294747Z digest=sha256:05acde4351e6764cc3b7e8fd4a6e17440863309de220a2dda92fe39129c52cbb

Observation 76779232-788b-4f44-8ea5-5f673283996a · outbound

This paper cites Large lan- guage models are zero-shot reasoners,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Large lan- guage models are zero-shot reasoners,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.299866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.299866Z digest=sha256:a5f2f345113dd4d023d6b028a446b19493f32bdba70171a97bb8d161270044e4

Observation df9a7451-b3c2-4b4f-919b-4f0f479b29ab · outbound

This paper cites Language mod- els are few-shot learners,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Language mod- els are few-shot learners,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.305123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.305123Z digest=sha256:9b28625f3b364614a9d4c2e1d743155a88c2d4821bbcd707dabb202d1438659e

Observation 4cfeeb3d-d209-4133-bc10-1c835bd9c235 · outbound

This paper cites A new measure of rank correlation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension A new measure of rank correlation,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.310133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.310133Z digest=sha256:24c493a2b7e07d346389ce8136961387615a1702d1124b6d4b929e16180108a4

Observation 58cd1b00-3a99-4ab5-b1d2-a38d1dd020a7 · outbound

This paper cites Pearson correlation coefficient,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Pearson correlation coefficient,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.315250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.315250Z digest=sha256:7687f2d1680d8264cea475d794dcf3a344a6fcd782146f61546f3559bf57a3bd

Observation 31a71bb3-22ba-4f7b-ba02-01f535efb7b9 · outbound

This paper cites Likert scale: Explored and explained,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Likert scale: Explored and explained,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.321002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.321002Z digest=sha256:3c87a9cf1b0aa3f43fd76693ed66734a63ca2d651197e0e92cfb4274ab7e2a4a

Observation 65067bc0-9be7-446d-9b5c-1c0d2785c538 · outbound

This paper cites Unsu- pervised translation of programming languages,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Unsu- pervised translation of programming languages,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.641453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.326026Z digest=sha256:c09b1f0097d365d3744e665e236b5effdfec29e4f0278c75fd1e4a1b52023930

Observation 98cca902-aa89-44ac-80d1-ee826433265b · outbound

This paper cites An empirical study of auto- mated unit test generation for python,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension An empirical study of auto- mated unit test generation for python,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.623676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T05:35:51.331017Z digest=sha256:7a68b7be818f289c731feb862d59551eb6eb8c1887625bf4da73568991dcbd4d

Observation 24059aa6-2376-47a7-9ebe-1ebf7aeeb741 · outbound

This paper cites Can Large Language Models Be an Alternative to Human Evaluations?.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Can Large Language Models Be an Alternative to Human Evaluations?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.335834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.335834Z digest=sha256:e31b8be70d80893de4b6cfa70474b719572e23d99499e481ad88e735a7a89fd2

Observation b4e5cb43-f826-4735-be77-653f9dbce629 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.340765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.340765Z digest=sha256:4efbb56ecaf432d0bff227963829afead60d48b495470cb94429c5284e50a670

Pith citing papers

Observation 107c2bc4-bd65-498b-8a61-d813d5839c35 · inbound

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection cites this paper.

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:24:28.422461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T16:24:28.140998Z digest=sha256:feff122d2139d6f83a192bbfcd33ddeb1cf91d77579ff6673b767c215fffb705