Pith. sign in

Paper Citation Record · LEDGER

Quality Assessment of Python Tests Generated by Large Language Models

As of 7 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2506.14297.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14297 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:22:57.229784Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T17:37:51.790000Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:46:10.824386Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact5
  • verified fuzzy3
  • unresolved28
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2ce75d2-4b3d-4acf-8427-82c68f5519f3 · outbound

This paper cites Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models.

Quality Assessment of Python Tests Generated by Large Language Models Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:23:00.522246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:52.907011Z digest=sha256:2e88d9c479e3c29633eea0dbf1e7c333f7347607047deed97711f3708e33b635

Observation 2baf4883-dc09-4d31-ad66-46f3a5a069e6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:52.956816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:52.956816Z digest=sha256:b9b203ca142f8459ca8a8b5eabd1ce9e76a57276ad174054ffd748dd1ae967b4

Observation b74754e8-2605-493f-add0-a45212bee263 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:23:00.320326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:53.166885Z digest=sha256:e34a9eb4b7cf705da689949c55461fb6c8450d8efaa1f60a085696a31cf00342

Observation c992e3e7-edcd-42eb-aa68-d3d4ea0763e0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 5

Resolution
malformed identifier
no resolver link, observed 2026-08-07T00:22:53.383429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.383429Z digest=sha256:914835a551857e1d7321f27fdd5b59fc0825b66379c776c8cae09c80d62b48a4

Observation a591b136-a682-4b39-8399-501e867fa2c6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T00:22:57.527939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:53.465341Z digest=sha256:632c304224289c467578f5a44452a60dc18976c18a8627f4659cc93f2ad637e7

Observation 3492489c-369f-471f-aa77-e76fca365e09 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating Large Language Models Trained on Code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.528702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.528702Z digest=sha256:4b7b9435142f363b2e7b3b0f87b24f3e4c17a68d4c711ca9a858e1da9199967b

Observation 7c700471-b715-4be9-932c-072a26a67464 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.620863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.620863Z digest=sha256:eaca424d096d289e4a5211f1f05507613fb3bdcb323bee057e20d6ed8d8d67f1

Observation f53d4363-7104-4d7a-be28-8c88d3f2c9e3 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.777526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.777526Z digest=sha256:43d7b3aca51a71ad94da6316f336824c32a3595f93a8ba86e8c5cc49690894bd

Observation 3dbac20d-5c75-4c0f-86c4-021eaf678633 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:59.852913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:53.861129Z digest=sha256:10eb88bf95ad1a25c9cd657a47eac19b2c675c7d8a6ef5348244b8476a61954e

Observation 05025f6e-0f17-4f8f-9cf2-b90afd9365ab · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 12

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:59.627204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:53.926091Z digest=sha256:eacd2378b3f61d961c16f1a845c5983f1307cad80f67f22787f4b5a6c96f801d

Observation f0773172-c8a0-4b84-ab6c-cc268ddd6c2d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.000232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.000232Z digest=sha256:868f3088982ee5efdc590434d3bdedc9d402372c3b0ed9aacb4997ef8c194e0f

Observation 6f2a4774-4e72-495a-88aa-4d5328114048 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.048447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.048447Z digest=sha256:c0167021f9bdbfe89071f5472a93ffd316e0160e8cc0da73dc25da5f3c64d0da

Observation 406b2339-ec23-49fb-ae9d-635d5e903e2f · outbound

This paper cites Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan.

Quality Assessment of Python Tests Generated by Large Language Models Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.117566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.117566Z digest=sha256:ae49504716bb294b81e3aeee7d9919c0ddc3d297dde1415ea479a9fd08e9b0e0

Observation 652449a9-dd3d-4787-b770-88486c2f2ec8 · outbound

This paper cites Graham, R.

Quality Assessment of Python Tests Generated by Large Language Models Graham, R

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.553263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:54.188448Z digest=sha256:001129e22b50038eb5f746f476f1829eb30d7e6acd692e07977c945bd8d84142

Observation 4d2b26f3-3ec7-4b32-8016-91961208dc66 · outbound

This paper cites 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot.

Quality Assessment of Python Tests Generated by Large Language Models 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.495272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:54.285734Z digest=sha256:1473d8d2ce6a46e49a9afef47d2a67079ab43cf2a0fc002e5fc364b975ac96b8

Observation 714c059f-df9c-4ebc-82a6-8e96fafcbce3 · outbound

This paper cites Khorikov.

Quality Assessment of Python Tests Generated by Large Language Models Khorikov

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:54.290955Z digest=sha256:1a5b93e6aa99bbbc471da8ed21011ba22fafdc2c679ff545f38f48439331ac02

Observation 30ff293e-bddf-4993-aa49-e27b5f4f3b5f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.267075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:54.366203Z digest=sha256:f8035be53dada637c6832c71aef2c7ee3948b380084b6856af7c1a05aec1e98b

Observation 733ac602-bcdc-44aa-8cb6-c23bccdb624f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.510354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.510354Z digest=sha256:656de2666c56723a2e90f1afc7495c629f9b97df2f153f772bbf8c97c904e006

Observation 7b5b833e-b19b-4dc0-a875-998f9e59c5a5 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.593557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.593557Z digest=sha256:6328b9c21110755ce4abee388451dbc3d3521e0a547285e1279f83c58b146fbb

Observation e591b293-72b1-4234-9c51-0df5c84c48a8 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

Quality Assessment of Python Tests Generated by Large Language Models CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.681694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.681694Z digest=sha256:209a7ea80d95e5228c55b78ad656289d684522ccd4c51f4155e84092221a6928

Observation ac2fe842-4109-4667-8115-26f6bad1abd4 · outbound

This paper cites Search-based software test data generation using evolutionary computation.

Quality Assessment of Python Tests Generated by Large Language Models Search-based software test data generation using evolutionary computation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:22:59.126324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:54.843366Z digest=sha256:024788e494073a0d3e5b69e63fe4cc5d55feab05168155c7a24afcfdfa3f3c16

Observation d7e546bd-95d7-4e47-9ef4-c4df30528740 · outbound

This paper cites Marvin, N.

Quality Assessment of Python Tests Generated by Large Language Models Marvin, N

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.900003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.900003Z digest=sha256:3ccef674486dd51101a403a040dd19c9a07e1e298d6e8179c5ea3cb1f3965d43

Observation 7f82476d-740b-4305-80da-8fb4cd17db78 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.190255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:55.164951Z digest=sha256:ae60b1d1326f676d331a776805b0b635a708936f0f596ac2a30fa0103a5a4e95

Observation 2eae1305-74c4-4636-a1fb-1b625e8ab18a · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 27

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:58.832867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:55.292198Z digest=sha256:15c241a54e507dbf7347124d4c0037036f69ca5589325ae4a5ce3ae6059c79ee

Observation 11045527-012f-422f-8f98-653f9768ffc2 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.412598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.412598Z digest=sha256:b44091831329f7568fc1edc4cac815aa61e88cdf7c2985a7fe1f80850dc13752

Observation 45045dd9-da6f-4acb-b9dd-f5b383544d3f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.499037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.499037Z digest=sha256:cfcc621cfb0d43aec0b54659efa8651088cf2615eb45bff2e1e49e1ed3b1b726

Observation ef0d7043-afe7-4524-90f7-08061ce8e248 · outbound

This paper cites Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen.

Quality Assessment of Python Tests Generated by Large Language Models Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.606582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.606582Z digest=sha256:c7aed6255c2cbbfa5cea42db7b49da8fdeabe57ce50c1e614e9c1be4c7085c34

Observation d8e185d7-c9ad-4032-97ed-2d9d2b5b99b6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.733894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.733894Z digest=sha256:b96353f6fe68cbfea93261788ab9d06931c23ff3b82701e7109fa968a7408b02

Observation b5459b6b-81ea-482d-94b1-af2126f754a0 · outbound

This paper cites A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications.

Quality Assessment of Python Tests Generated by Large Language Models A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.803426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.803426Z digest=sha256:e2f2cfc89acc5d3039a571535cb458fdc8e2e566b50f4a093b5ec38670586b14

Observation 43a2ce56-4b73-4c8c-b842-8b7dc576f190 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.023171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:55.889640Z digest=sha256:f114426f0de1322d380ea9c39801bb1898535d83b60022c0acc96e149c4cb3e0

Observation 8b96de58-5399-4419-ad44-ca45813a2857 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.959286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.959286Z digest=sha256:28b0005ebf8be5333f37e5130dc2509ed6e685e907e8647f0dca3959bf4aa418

Observation b7803329-62f0-4a4b-871e-aff5426195d1 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 35

Resolution
malformed identifier
doi_truncated, observed 2026-08-07T00:22:58.159750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:56.099483Z digest=sha256:480621143f0595ecb90722cdfe736929b2b249783fe420aeda6969ce3d0a4265

Observation f630d06a-d59c-4f07-8e21-b1f445bbf636 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 36

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:23:00.129940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:56.206882Z digest=sha256:f830bcabd5f1599144f13fd8e067e024c575af74632bbb05426e0a0f84f1df4a

Observation a362e75d-eb7b-4d96-a0fe-f4c181f31073 · outbound

This paper cites Unit Test Case Generation with Transformers and Focal Context.

Quality Assessment of Python Tests Generated by Large Language Models Unit Test Case Generation with Transformers and Focal Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.344629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.344629Z digest=sha256:abf7897340e3c251c5977971c61020696ac8c01c98a8cde6d761dbc5e19fb12e

Observation b5f5786d-522c-4a62-917c-d3003a413ac0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.473597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.473597Z digest=sha256:b87f6d7a05d0b6d43c80ae3f7a9e5602d7d9529983c4c5db8fab0bc39773500a

Observation f45fb04c-3f11-425c-ac28-c4f22366833d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.923542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:56.564574Z digest=sha256:6faf5cd9da5a3d3ceea0a9924e19c18f6452ad9c60296fbdbc58d89bc565a916

Observation 8f52b362-fb2f-43c2-9899-2532c4bca62b · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 40

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:57.843768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:56.694578Z digest=sha256:bcbe586c69ac12455c350a47005b1294050fac716159d0ead55515b32212e0be

Observation 4a8b768d-a97c-4a77-8ef5-bcbe7c4bc016 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.791389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.791389Z digest=sha256:b250b56621fea0ace8d11690c886dcab5974866061f0e22e0bb4a82bf9cbbe6c

Observation 4a1f0b84-16fa-41e4-bf57-aae274d446a0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.781962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:56.914404Z digest=sha256:51bc5ad63574f4bb12796362990a5de0124adf35fe291a9eefc7c5f61e950ecd

Observation 20852e48-d589-4406-8d08-7d1439c0ea86 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.025241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.025241Z digest=sha256:69bf04437403ec60e5a896e119409e7b958e04c80d46db627609a258db9c0a86

Observation 9da3e49d-a12e-44d9-b41c-d7f49378169f · outbound

This paper cites Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.143131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.143131Z digest=sha256:04ad71438c8d3c6804653107f401a8a959035db8eff4eb4b3ef532080a9c22b5

Observation 32ddb553-d030-44a7-b15e-9ce263555191 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.629614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:22:57.229784Z digest=sha256:f4a09cae4363f49a5daed0b454782d73e2717952279a743cadbc94b96d7331c5

Pith citing papers

Observation 039d6874-8fbe-4269-9744-543facd57397 · inbound

Can LLMs be Effective Code Contributors? A Study on Open-source Projects cites this paper.

Can LLMs be Effective Code Contributors? A Study on Open-source Projects Quality Assessment of Python Tests Generated by Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:10.830944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T08:09:33.211692Z digest=sha256:4d1540742598a0600431d7da2b21e9fcb3cd382241422d7e36e42f52ca86bd94

Observation a220b4a9-3197-41ba-95b2-36e94fffb31a · inbound

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code cites this paper.

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code Quality Assessment of Python Tests Generated by Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:26:04.393629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T17:37:51.790000Z digest=sha256:4f7c195f63e5391c0a53505fce89364d82075e31f0f9d9bdeeec47b38f09b331