Pith. sign in

Paper Citation Record · LEDGER

From Code to Courtroom: LLMs as the New Software Judges

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2503.02246.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.02246 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:55:59.045171Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T09:56:10.709911Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ed4a21dc-a8ef-4e74-bb21-651ce0b3d667 · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models From Code to Courtroom: LLMs as the New Software Judges

Reference 247

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:40:41.499989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:9e2453695c9c26806f537ddcb846e241bb1a700d6e434eb496df09d04ac05b83

Observation e83c24d1-2422-4c50-9cdd-4cbdb8818293 · inbound

AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers cites this paper.

AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers From Code to Courtroom: LLMs as the New Software Judges

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T05:55:59.045171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:55:59.045171Z digest=sha256:5a2b94b60dce7222b3ca3f1a845f9a0e915cceeff1c649e7075521037a486f78

Observation 20add06a-1154-416b-8b05-f3a185bb4f3b · inbound

CODE-DITING: A Reasoning-Based Metric for Functional Alignment in Code Evaluation cites this paper.

CODE-DITING: A Reasoning-Based Metric for Functional Alignment in Code Evaluation From Code to Courtroom: LLMs as the New Software Judges

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:48.242132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:48.242132Z digest=sha256:6fc9cbee7a96cf0d64b91c259712f076107b72a782c6ff1e49dca891b705b3e2

Observation 9a4f0a3c-c7f6-404d-8446-79094505de63 · inbound

CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review cites this paper.

CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review From Code to Courtroom: LLMs as the New Software Judges

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:28.712917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:13:28.712917Z digest=sha256:8ce6cb5e33ad117a9b7a46266decf6369255835829516618511760cc28a90f09

Observation 2265ac91-5434-4ff7-934c-790046f0b79f · inbound

CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking cites this paper.

CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking From Code to Courtroom: LLMs as the New Software Judges

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:57:22.059743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:57:22.059743Z digest=sha256:044684911921871fb0358de833e24cddd011a2827b7ad670fd2177483c003771

Observation 6c19e092-2e61-4550-8117-b1a7ec4a4f62 · inbound

CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks cites this paper.

CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks From Code to Courtroom: LLMs as the New Software Judges

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:32:31.444375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:32:31.444375Z digest=sha256:c73013e86749f38381e354f99e009f696293e56f557b1f268a2015fbbea9a9b7

Observation e304bf91-6121-44ce-9b45-b8a4a3137f34 · inbound

SWE-QA: Can Language Models Answer Repository-level Code Questions? cites this paper.

SWE-QA: Can Language Models Answer Repository-level Code Questions? From Code to Courtroom: LLMs as the New Software Judges

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:41:37.755478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T16:41:29.829199Z digest=sha256:7e521b894892d6dbbb0c5ce930d678e9ba7fec5e86b12e1a9cfee4530546d0fa

Observation 993c43a8-10ed-41ae-96c9-43f4d61e4dc4 · inbound

CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases cites this paper.

CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases From Code to Courtroom: LLMs as the New Software Judges

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:10:48.440104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T03:10:21.188635Z digest=sha256:3ec69d5e1239a0067b77e5b95a1cc4baccc59973d9e77695e8006beea16dfe99

Observation 37b321e7-e9ba-4324-a1e3-b90f64704d0c · inbound

Bias in the Loop: Auditing LLM-as-a-Judge for Software Engineering cites this paper.

Bias in the Loop: Auditing LLM-as-a-Judge for Software Engineering From Code to Courtroom: LLMs as the New Software Judges

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:32:00.325130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T07:29:03.994957Z digest=sha256:9c2655c7dfc8d7ff02a4aaa56f27762a4d90c91d405331a013b064c1eee2f713

Observation 9ff977c7-63a7-477f-9d4c-ecb727133725 · inbound

An Empirical Study on Logging Evolution On Stack Overflow: Trends, Topics, and Challenges cites this paper.

An Empirical Study on Logging Evolution On Stack Overflow: Trends, Topics, and Challenges From Code to Courtroom: LLMs as the New Software Judges

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:33:18.393631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T10:32:47.756343Z digest=sha256:52d83dfe2309cae7a984498ead623032cb389a20ea17e4c9ffef96375550f9cc

Observation f3d01e4c-68e5-4ed0-9f85-a1949c880123 · inbound

Biased or Personalized? The Impact of Personal Information on AI-driven Development cites this paper.

Biased or Personalized? The Impact of Personal Information on AI-driven Development From Code to Courtroom: LLMs as the New Software Judges

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-09T09:56:10.712455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-09T09:46:37.332477Z digest=sha256:354ffbcf4d58e7e1e905af1558c3d816955acddb0b8b7c674983e924b26ab732

Observation aa826707-4463-443f-b1c5-862c18f01932 · inbound

Bridging Behavior and Implementation: Automated Java Glue Code Generation for Behavior-Driven Development cites this paper.

Bridging Behavior and Implementation: Automated Java Glue Code Generation for Behavior-Driven Development From Code to Courtroom: LLMs as the New Software Judges

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T12:00:24.380603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:00:24.380603Z digest=sha256:11fcde6fe8a8e83cf6dc08da12f3e19dc71a23555cd96e8c9d2dfe818c06d980

Observation 96f024b4-86c1-4da2-9ba4-7e90311fb13b · inbound

Detecting Soft Skills in ML Engineering Roles CVs cites this paper.

Detecting Soft Skills in ML Engineering Roles CVs From Code to Courtroom: LLMs as the New Software Judges

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:55.249273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:55.249273Z digest=sha256:24ece487aa55ee2819e5c5f2c74842255e99a3e0cc6eef90f227704bed079765