Pith. sign in

Paper Citation Record · LEDGER

Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2404.13340.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.13340 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:44:09.359830Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T08:37:11.574579Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 004312e9-39b0-40ad-848d-1db5e78ba150 · inbound

Language Models in Software Development Tasks: An Experimental Analysis of Energy and Accuracy cites this paper.

Language Models in Software Development Tasks: An Experimental Analysis of Energy and Accuracy Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T05:32:45.246532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:32:45.246532Z digest=sha256:1b3c0cf5f97e21d6874f81054daeb30b4f3c55379c12bba7646c81db936ac92a

Observation 4ffa50f1-9a19-4f62-8455-b5d2995fb53f · inbound

COFFE: A Code Efficiency Benchmark for Code Generation cites this paper.

COFFE: A Code Efficiency Benchmark for Code Generation Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T11:01:27.805110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:01:27.805110Z digest=sha256:ae2a14fadf59fa9d62ab8ccc254ac2f1cc7c062bd601f0372c9823fc95cb1e83

Observation a496c04a-024a-4b8a-b31b-901ba4379ed1 · inbound

Large Language Model Guided Self-Debugging Code Generation cites this paper.

Large Language Model Guided Self-Debugging Code Generation Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T10:38:44.960968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:38:44.960968Z digest=sha256:bde8f4c21edbfa0ed73c50ce06d0265f4561b68711307d3e89d602500770ed09

Observation fe8554fe-349f-4db5-b86d-8b5888d776f4 · inbound

Large Language Models for Unit Testing: A Systematic Literature Review cites this paper.

Large Language Models for Unit Testing: A Systematic Literature Review Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T19:44:09.359830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:44:09.359830Z digest=sha256:41eba0a6e9aab45c872c9acd38eebe2a2f5d8a5724cde554bbf02b7ef6d27085

Observation 84e8d582-73e8-49ab-9281-35ed10768ea8 · inbound

Effective LLM Code Refinement via Property-Oriented and Structurally Minimal Feedback cites this paper.

Effective LLM Code Refinement via Property-Oriented and Structurally Minimal Feedback Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:37:11.580811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-19T08:36:58.880345Z digest=sha256:b96dbbbc2cbd595ab20c0ed4443f954df3e2efc6d9624cda8db4370bdb06237e

Observation 27e392bd-6d9c-43fc-b7d5-f51147e01cf1 · inbound

Rethinking Verification for LLM Code Generation: From Generation to Testing cites this paper.

Rethinking Verification for LLM Code Generation: From Generation to Testing Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:58:21.120298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:58:21.120298Z digest=sha256:1385c057d4ce636f6fa38ed3555154f48eb8de098895a68630305362a979c300

Observation 94f29f9c-56b4-4ba3-8861-6c63c4d4091c · inbound

HyClone: Bridging LLM Understanding and Dynamic Execution for Semantic Code Clone Detection cites this paper.

HyClone: Bridging LLM Understanding and Dynamic Execution for Semantic Code Clone Detection Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T05:41:35.095742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:41:35.095742Z digest=sha256:1fdf09e0ba61208717379115702746eeb68a265b41af5cdc3e08c496eea6a22a

Observation 2b57733b-3198-454b-89b0-8188e84b0a0d · inbound

Trade Policy and Structural Change cites this paper.

Trade Policy and Structural Change Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T05:42:30.551364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:42:30.551364Z digest=sha256:cb61f959e071e5c89c6a8ccaf08e1a32eaf4c0a5d007b3b9f6604034c8473d82

Observation ae68aeb3-b38a-45e3-8125-6486bb02546b · inbound

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation cites this paper.

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:53:14.033218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-14T22:52:58.524934Z digest=sha256:cdec867e1a47d42c44c85e4617c3401ea33a6be9690ca87544e987bf16bc1f20

Observation 70fd33f0-7e7c-42a9-a47b-28801c219b98 · inbound

Enhancing Large Language Models with Retrieval Augmented Generation for Software Testing and Inspection Automation cites this paper.

Enhancing Large Language Models with Retrieval Augmented Generation for Software Testing and Inspection Automation Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:44:38.002962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T10:40:30.734804Z digest=sha256:7b4201a6aa44e430c9ff7287aa3a9a38354bdd24d54c4a7f7f3d935f58a5099d

Observation 63165519-3557-49cc-a62f-8dd91d183266 · inbound

MineValiCoder: Reliable Code Generation with Test Case Quality Mining and Bipartite Graph-Based Mutual Validation cites this paper.

MineValiCoder: Reliable Code Generation with Test Case Quality Mining and Bipartite Graph-Based Mutual Validation Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T04:45:11.001231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:45:11.001231Z digest=sha256:295d73fdb6d313b209286af3eef0d1b283394263663ff1fc233405bcea73c966

Observation e97dca2f-fea2-4ae6-bb99-c3148d2eab12 · inbound

PROGRESS: Property-Guided Regression Search for Semantic Falsification cites this paper.

PROGRESS: Property-Guided Regression Search for Semantic Falsification Large Language Models as Test Case Generators: Performance Evaluation and Enhancement

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T08:56:26.123848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:56:26.123848Z digest=sha256:ccee74f9a30adef6e221c55ba96445bfa7cfe99b437938204e2809f102cb0219