Pith. sign in

Paper Citation Record · LEDGER

Quality Assessment of Python Tests Generated by Large Language Models

As of 17 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2506.14297.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14297 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:22:57.229784Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T17:37:51.790000Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:46:10.824386Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact5
  • verified fuzzy3
  • unresolved28
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2ce75d2-4b3d-4acf-8427-82c68f5519f3 · outbound

This paper cites Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models.

Quality Assessment of Python Tests Generated by Large Language Models Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:23:00.522246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:52.907011Z digest=sha256:9c515b72e1ec29d2c4d913e42375c302d565e3a060d6fe1b2d6d733c06286a15

Observation 2baf4883-dc09-4d31-ad66-46f3a5a069e6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:52.956816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:52.956816Z digest=sha256:7f6665723b1412878a06edf46f48ee1ee4ff631ab96b6dddd710b30b50cbf650

Observation b74754e8-2605-493f-add0-a45212bee263 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:23:00.320326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:53.166885Z digest=sha256:fce1825526aa24830fb1accd4dfabc301f884dcd595ab31e962803d350d24503

Observation c992e3e7-edcd-42eb-aa68-d3d4ea0763e0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 5

Resolution
malformed identifier
no resolver link, observed 2026-08-07T00:22:53.383429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.383429Z digest=sha256:6911e9e95c9fc80ec659bdcfbe861694699bf6bf3c743b7d70c18e4b2494bcaa

Observation a591b136-a682-4b39-8399-501e867fa2c6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T00:22:57.527939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:53.465341Z digest=sha256:aa0b73a41ab85b9b7488690a6a67e5dbb020eb650f7243a16afede77a1544e41

Observation 3492489c-369f-471f-aa77-e76fca365e09 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating Large Language Models Trained on Code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.528702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.528702Z digest=sha256:f85e2e8679b04f82e9a5ec1991b2aaaf4637d30c8097590e9171ec4a71c26d0a

Observation 7c700471-b715-4be9-932c-072a26a67464 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.620863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.620863Z digest=sha256:d578a8e1e92acc11dfa7a86c47038bd5cb8de0c90f5fb95dbe37a7f874ea9af3

Observation f53d4363-7104-4d7a-be28-8c88d3f2c9e3 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.777526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.777526Z digest=sha256:983579f254a47a98ee802caa137b19db552c71094c1dd50ec52b93959beedd2a

Observation 3dbac20d-5c75-4c0f-86c4-021eaf678633 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:59.852913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:53.861129Z digest=sha256:60060ec61b9acd3a02b33225cf9751342214b64f7c2ed14b99f94f1950c77e5c

Observation 05025f6e-0f17-4f8f-9cf2-b90afd9365ab · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 12

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:59.627204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:53.926091Z digest=sha256:c5869ff169cfc0f51559bfce5d4f109d2ed3c2919899f4746753dff9459c6db6

Observation f0773172-c8a0-4b84-ab6c-cc268ddd6c2d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.000232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.000232Z digest=sha256:46aa65f8826161e70c18756b0b1726ee0e8a2337133703890ae20c261fd44d73

Observation 6f2a4774-4e72-495a-88aa-4d5328114048 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.048447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.048447Z digest=sha256:a3c73479edf4cc6d786bdb8055f1eb185e75cb5fce5cfd546e71c9d8039506f7

Observation 406b2339-ec23-49fb-ae9d-635d5e903e2f · outbound

This paper cites Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan.

Quality Assessment of Python Tests Generated by Large Language Models Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.117566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.117566Z digest=sha256:6656f0151cf5a2cab1219f2850e21d86860682785251b68814cb699c2c2ea38b

Observation 652449a9-dd3d-4787-b770-88486c2f2ec8 · outbound

This paper cites Graham, R.

Quality Assessment of Python Tests Generated by Large Language Models Graham, R

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.553263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:54.188448Z digest=sha256:29912ed568ea672c409cd1e09886ba59d0db02b80d75b4fbedad782624207a97

Observation 4d2b26f3-3ec7-4b32-8016-91961208dc66 · outbound

This paper cites 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot.

Quality Assessment of Python Tests Generated by Large Language Models 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.495272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:54.285734Z digest=sha256:51eb7fc0ae21d250f46663c23a3ce633120221f2561a55f868ec7adfb680dd96

Observation 714c059f-df9c-4ebc-82a6-8e96fafcbce3 · outbound

This paper cites Khorikov.

Quality Assessment of Python Tests Generated by Large Language Models Khorikov

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:54.290955Z digest=sha256:7192c4ff5560601a818f2521a291bfc8d4e33299789b6f393bcc9759a7de227b

Observation 30ff293e-bddf-4993-aa49-e27b5f4f3b5f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.267075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:54.366203Z digest=sha256:84d5f28699b06e1736f6f96eace5e20a15653d62a4c5170ceb541354d673123e

Observation 733ac602-bcdc-44aa-8cb6-c23bccdb624f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.510354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.510354Z digest=sha256:aa00fd460bb0284f7f4f33a1c3a203e0a32b158bbf76f54a798800b4ff6fae08

Observation 7b5b833e-b19b-4dc0-a875-998f9e59c5a5 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.593557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.593557Z digest=sha256:808b63704970c9002a0bd24eee41fd84f4ead5ea409c952008728431ef78e6a8

Observation e591b293-72b1-4234-9c51-0df5c84c48a8 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

Quality Assessment of Python Tests Generated by Large Language Models CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.681694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.681694Z digest=sha256:43875767e0be4489c505c3aa79ebf3cbfa818dcd1cb982ca87f151e458483365

Observation ac2fe842-4109-4667-8115-26f6bad1abd4 · outbound

This paper cites Search-based software test data generation using evolutionary computation.

Quality Assessment of Python Tests Generated by Large Language Models Search-based software test data generation using evolutionary computation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:22:59.126324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:54.843366Z digest=sha256:8ce63eb48af197946cdd4b5ca72625b4d732a4f58cbb86ab6ed3c9caa8f4721f

Observation d7e546bd-95d7-4e47-9ef4-c4df30528740 · outbound

This paper cites Marvin, N.

Quality Assessment of Python Tests Generated by Large Language Models Marvin, N

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.900003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.900003Z digest=sha256:b890b551e0adab79606806063c29f2f759a82df386dcc6e089c20188c4a5a813

Observation 7f82476d-740b-4305-80da-8fb4cd17db78 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.190255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:55.164951Z digest=sha256:a13eeb489c20a1241a09091abdeee3f6cb809a7a5fe721eb719dcef7a9560266

Observation 2eae1305-74c4-4636-a1fb-1b625e8ab18a · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 27

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:58.832867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:55.292198Z digest=sha256:8e100eb0c77e0480d7e62bb39ff0c66b4b4e49cde14d3018541ef67751090f77

Observation 11045527-012f-422f-8f98-653f9768ffc2 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.412598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.412598Z digest=sha256:70154e050496e6b4f854ec7013d14d385da2879bb203dec3cf25a0b7292cd624

Observation 45045dd9-da6f-4acb-b9dd-f5b383544d3f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.499037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.499037Z digest=sha256:72b90de3396c941c4022307065777f328e5f795fffe040c7bc6c421bd5c7e170

Observation ef0d7043-afe7-4524-90f7-08061ce8e248 · outbound

This paper cites Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen.

Quality Assessment of Python Tests Generated by Large Language Models Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.606582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.606582Z digest=sha256:6ca2b93f93343ff6d8ed204488de35eb422a9645a748937e00cd043eee379a2a

Observation d8e185d7-c9ad-4032-97ed-2d9d2b5b99b6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.733894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.733894Z digest=sha256:005cf34ac8402a00f219a019846cd95e4652bb10e1f3450860991e11dc103723

Observation b5459b6b-81ea-482d-94b1-af2126f754a0 · outbound

This paper cites A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications.

Quality Assessment of Python Tests Generated by Large Language Models A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.803426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.803426Z digest=sha256:741b7440e7b8f5d5894692a392eaf8067c1a1b4bc71c6f75e0318bbfefa4f2e2

Observation 43a2ce56-4b73-4c8c-b842-8b7dc576f190 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.023171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:55.889640Z digest=sha256:fa67789cb7ec6693aa504f14e57580087ba5f82be89923a144a678e9e6c26d9f

Observation 8b96de58-5399-4419-ad44-ca45813a2857 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.959286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.959286Z digest=sha256:23e4a9cb24ef26e88fd330c56d00e41ab3079f5aad76f445ab9545297f3f1166

Observation b7803329-62f0-4a4b-871e-aff5426195d1 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 35

Resolution
malformed identifier
doi_truncated, observed 2026-08-07T00:22:58.159750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:56.099483Z digest=sha256:e4d0b572e846ef528f37ec57c736f7de9efbf5d643f6a30935ef505d2a8995fa

Observation f630d06a-d59c-4f07-8e21-b1f445bbf636 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 36

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:23:00.129940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:56.206882Z digest=sha256:659e13c7ebcea506713656436ca4364e3bcbbee355230eb707134df05a4f61f2

Observation a362e75d-eb7b-4d96-a0fe-f4c181f31073 · outbound

This paper cites Unit Test Case Generation with Transformers and Focal Context.

Quality Assessment of Python Tests Generated by Large Language Models Unit Test Case Generation with Transformers and Focal Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.344629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.344629Z digest=sha256:84bd3e697c96b988d3ed6c9183f429240300f8cabe4098d34e8cb47ca65ae412

Observation b5f5786d-522c-4a62-917c-d3003a413ac0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.473597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.473597Z digest=sha256:32ba6e871a4aba3383f574686768e1d1f01f0f538055c8321263617bcbb36bf2

Observation f45fb04c-3f11-425c-ac28-c4f22366833d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.923542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:56.564574Z digest=sha256:f7866c24c150ef525f0827e7152c653d6cd3ef6db1bc657564034133f44f17a8

Observation 8f52b362-fb2f-43c2-9899-2532c4bca62b · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 40

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:57.843768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:56.694578Z digest=sha256:545c0ce46c48b0c20426433162e5c9a9c937795eca9bcc2893330e144c9a1440

Observation 4a8b768d-a97c-4a77-8ef5-bcbe7c4bc016 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.791389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.791389Z digest=sha256:69eb123cb1a81dde5548a07d9b0ea79ae34bddf3471103b629cb66fa211c70a8

Observation 4a1f0b84-16fa-41e4-bf57-aae274d446a0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.781962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:56.914404Z digest=sha256:1622e31bbb8e4d550b11b02cb2588db312668ad52c5efc685418d236265acc14

Observation 20852e48-d589-4406-8d08-7d1439c0ea86 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.025241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.025241Z digest=sha256:9cb282a6520ddc517a640c59e25a1386c23b1b911836ac4dc7fe332a885cb76f

Observation 9da3e49d-a12e-44d9-b41c-d7f49378169f · outbound

This paper cites Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.143131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.143131Z digest=sha256:41dd8e004f38c9130ca3691baaaa556796d3673a1cc2b60a754fa4658a02d4f8

Observation 32ddb553-d030-44a7-b15e-9ce263555191 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.629614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:22:57.229784Z digest=sha256:287381e41c5e13d29e764a8791dca32fd8a3d46766ca6e01edff992b38583206

Pith citing papers

Observation 039d6874-8fbe-4269-9744-543facd57397 · inbound

Can LLMs be Effective Code Contributors? A Study on Open-source Projects cites this paper.

Can LLMs be Effective Code Contributors? A Study on Open-source Projects Quality Assessment of Python Tests Generated by Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:10.830944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T08:09:33.211692Z digest=sha256:a3e8def2624a6697310a773bab52db283f905c05d268da269d9073dad64d0e61

Observation a220b4a9-3197-41ba-95b2-36e94fffb31a · inbound

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code cites this paper.

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code Quality Assessment of Python Tests Generated by Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:26:04.393629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T17:37:51.790000Z digest=sha256:1dc2ec883667aa08fc4f5f909b37d48dc215dd1078444a4cea8125284dd16cf0