Pith. sign in

Paper Citation Record · LEDGER

Quality Assessment of Python Tests Generated by Large Language Models

As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2506.14297.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14297 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:22:57.229784Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T17:37:51.790000Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:46:10.824386Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact5
  • verified fuzzy3
  • unresolved28
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2ce75d2-4b3d-4acf-8427-82c68f5519f3 · outbound

This paper cites Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models.

Quality Assessment of Python Tests Generated by Large Language Models Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:23:00.522246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:52.907011Z digest=sha256:18c8c0e9d69b7929c5fe6c16b507951765855eaff42caffba11a6cf3872ebb5e

Observation 2baf4883-dc09-4d31-ad66-46f3a5a069e6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:52.956816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:52.956816Z digest=sha256:b9b203ca142f8459ca8a8b5eabd1ce9e76a57276ad174054ffd748dd1ae967b4

Observation b74754e8-2605-493f-add0-a45212bee263 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:23:00.320326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:53.166885Z digest=sha256:1491e8fc31976483452a359c7878b22dadd8086eb0dd54590e8794516c8575d7

Observation c992e3e7-edcd-42eb-aa68-d3d4ea0763e0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 5

Resolution
malformed identifier
no resolver link, observed 2026-08-07T00:22:53.383429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.383429Z digest=sha256:914835a551857e1d7321f27fdd5b59fc0825b66379c776c8cae09c80d62b48a4

Observation a591b136-a682-4b39-8399-501e867fa2c6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T00:22:57.527939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:53.465341Z digest=sha256:fdfe214a6578726075aa92e987c2e046739a3d36f06c9a09fef22a763e289327

Observation 3492489c-369f-471f-aa77-e76fca365e09 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating Large Language Models Trained on Code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.528702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.528702Z digest=sha256:c1e9bf5851aee55c26b7f98ece9aeff762e9b3c45d19f0d49059408bf845c523

Observation 7c700471-b715-4be9-932c-072a26a67464 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.620863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.620863Z digest=sha256:eaca424d096d289e4a5211f1f05507613fb3bdcb323bee057e20d6ed8d8d67f1

Observation f53d4363-7104-4d7a-be28-8c88d3f2c9e3 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:53.777526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:53.777526Z digest=sha256:43d7b3aca51a71ad94da6316f336824c32a3595f93a8ba86e8c5cc49690894bd

Observation 3dbac20d-5c75-4c0f-86c4-021eaf678633 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:59.852913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:53.861129Z digest=sha256:84da89ce0169cc70ee15d35094f7bdd7d055dd7ad4bbf755e5ba09815a31ab03

Observation 05025f6e-0f17-4f8f-9cf2-b90afd9365ab · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 12

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:59.627204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:53.926091Z digest=sha256:ecf973a98934c4be3d3ed259b8976712544ac2308ddde0a84a88ae0bd0f628ee

Observation f0773172-c8a0-4b84-ab6c-cc268ddd6c2d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.000232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.000232Z digest=sha256:868f3088982ee5efdc590434d3bdedc9d402372c3b0ed9aacb4997ef8c194e0f

Observation 6f2a4774-4e72-495a-88aa-4d5328114048 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.048447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.048447Z digest=sha256:c0167021f9bdbfe89071f5472a93ffd316e0160e8cc0da73dc25da5f3c64d0da

Observation 406b2339-ec23-49fb-ae9d-635d5e903e2f · outbound

This paper cites Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan.

Quality Assessment of Python Tests Generated by Large Language Models Santos, Andrew Popovich, Mehdi Mirakhorli, and Mei Nagappan

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.117566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.117566Z digest=sha256:ae49504716bb294b81e3aeee7d9919c0ddc3d297dde1415ea479a9fd08e9b0e0

Observation 652449a9-dd3d-4787-b770-88486c2f2ec8 · outbound

This paper cites Graham, R.

Quality Assessment of Python Tests Generated by Large Language Models Graham, R

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.553263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:54.188448Z digest=sha256:534c51a0a3a154de107e0bc35e115c78a3f7f361fc71e1242d70b65a8e8e8a0f

Observation 4d2b26f3-3ec7-4b32-8016-91961208dc66 · outbound

This paper cites 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot.

Quality Assessment of Python Tests Generated by Large Language Models 2023.Code Correctness and Quality in the Era of AI Code Generation: Examining ChatGPT and GitHub Copilot

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.495272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:54.285734Z digest=sha256:727673e235ce2afc2d1c5870028e8a3c22ad92c7d5432048090aae67337084bb

Observation 714c059f-df9c-4ebc-82a6-8e96fafcbce3 · outbound

This paper cites Khorikov.

Quality Assessment of Python Tests Generated by Large Language Models Khorikov

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:23:01.388956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:54.290955Z digest=sha256:274500db6e37b12bc4e858ef3570510a964405b433554f7e1e88c22327a93e66

Observation 30ff293e-bddf-4993-aa49-e27b5f4f3b5f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.267075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:54.366203Z digest=sha256:51169cdea701aabcbb5132a03b2ebe5d01bc795cad82a7dd372f36cd590f84a4

Observation 733ac602-bcdc-44aa-8cb6-c23bccdb624f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.510354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.510354Z digest=sha256:656de2666c56723a2e90f1afc7495c629f9b97df2f153f772bbf8c97c904e006

Observation 7b5b833e-b19b-4dc0-a875-998f9e59c5a5 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.593557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.593557Z digest=sha256:6328b9c21110755ce4abee388451dbc3d3521e0a547285e1279f83c58b146fbb

Observation e591b293-72b1-4234-9c51-0df5c84c48a8 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

Quality Assessment of Python Tests Generated by Large Language Models CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.681694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.681694Z digest=sha256:209a7ea80d95e5228c55b78ad656289d684522ccd4c51f4155e84092221a6928

Observation ac2fe842-4109-4667-8115-26f6bad1abd4 · outbound

This paper cites Search-based software test data generation using evolutionary computation.

Quality Assessment of Python Tests Generated by Large Language Models Search-based software test data generation using evolutionary computation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:22:59.126324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:54.843366Z digest=sha256:28e36a435e4e1a0918926f0e86f51f01d4b0ad1f53411f405d7cf1a007a2cb8e

Observation d7e546bd-95d7-4e47-9ef4-c4df30528740 · outbound

This paper cites Marvin, N.

Quality Assessment of Python Tests Generated by Large Language Models Marvin, N

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:54.900003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:54.900003Z digest=sha256:3ccef674486dd51101a403a040dd19c9a07e1e298d6e8179c5ea3cb1f3965d43

Observation 7f82476d-740b-4305-80da-8fb4cd17db78 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.190255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:55.164951Z digest=sha256:7e3b2650a7c84d91afcffe929c5ebdb30d62fea8c882192a84403df10b5054e0

Observation 2eae1305-74c4-4636-a1fb-1b625e8ab18a · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 27

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:22:58.832867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:55.292198Z digest=sha256:17c32586a0befb913a1328f7b40b55cec425220b0efd8140403785a5ba1ac7e3

Observation 11045527-012f-422f-8f98-653f9768ffc2 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.412598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.412598Z digest=sha256:b44091831329f7568fc1edc4cac815aa61e88cdf7c2985a7fe1f80850dc13752

Observation 45045dd9-da6f-4acb-b9dd-f5b383544d3f · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.499037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.499037Z digest=sha256:cfcc621cfb0d43aec0b54659efa8651088cf2615eb45bff2e1e49e1ed3b1b726

Observation ef0d7043-afe7-4524-90f7-08061ce8e248 · outbound

This paper cites Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen.

Quality Assessment of Python Tests Generated by Large Language Models Becker, Arto Hellas, Bailey Kimmel, Garrett Powell, and Juho Leinonen

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.606582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.606582Z digest=sha256:c7aed6255c2cbbfa5cea42db7b49da8fdeabe57ce50c1e614e9c1be4c7085c34

Observation d8e185d7-c9ad-4032-97ed-2d9d2b5b99b6 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.733894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.733894Z digest=sha256:b96353f6fe68cbfea93261788ab9d06931c23ff3b82701e7109fa968a7408b02

Observation b5459b6b-81ea-482d-94b1-af2126f754a0 · outbound

This paper cites A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications.

Quality Assessment of Python Tests Generated by Large Language Models A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.803426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.803426Z digest=sha256:e2f2cfc89acc5d3039a571535cb458fdc8e2e566b50f4a093b5ec38670586b14

Observation 43a2ce56-4b73-4c8c-b842-8b7dc576f190 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:01.023171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:55.889640Z digest=sha256:311113b1ef5a02bbce8c730e4e9f371e23c39b4e1f60f49c04546099957f620b

Observation 8b96de58-5399-4419-ad44-ca45813a2857 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:55.959286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:55.959286Z digest=sha256:28b0005ebf8be5333f37e5130dc2509ed6e685e907e8647f0dca3959bf4aa418

Observation b7803329-62f0-4a4b-871e-aff5426195d1 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 35

Resolution
malformed identifier
doi_truncated, observed 2026-08-07T00:22:58.159750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:56.099483Z digest=sha256:bdb522db90acdee46c093fe6d228caaa81a21414862a305aada1713b226f0d20

Observation f630d06a-d59c-4f07-8e21-b1f445bbf636 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 36

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:23:00.129940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:56.206882Z digest=sha256:780bc6d8020591cc87eebecd09fccf258888f51f79882d742dbdd1943827db27

Observation a362e75d-eb7b-4d96-a0fe-f4c181f31073 · outbound

This paper cites Unit Test Case Generation with Transformers and Focal Context.

Quality Assessment of Python Tests Generated by Large Language Models Unit Test Case Generation with Transformers and Focal Context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.344629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.344629Z digest=sha256:abf7897340e3c251c5977971c61020696ac8c01c98a8cde6d761dbc5e19fb12e

Observation b5f5786d-522c-4a62-917c-d3003a413ac0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.473597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.473597Z digest=sha256:b87f6d7a05d0b6d43c80ae3f7a9e5602d7d9529983c4c5db8fab0bc39773500a

Observation f45fb04c-3f11-425c-ac28-c4f22366833d · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.923542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:56.564574Z digest=sha256:05b5549d4fd0c850ff6736a6991b05470ffbf7ca37566d12b06bd22189cfd42e

Observation 8f52b362-fb2f-43c2-9899-2532c4bca62b · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 40

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:22:57.843768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:56.694578Z digest=sha256:b22fb19398e69bdff18a40236dcc807fa89d0b825bea1775092250d9d7e4856a

Observation 4a8b768d-a97c-4a77-8ef5-bcbe7c4bc016 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:56.791389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:56.791389Z digest=sha256:b250b56621fea0ace8d11690c886dcab5974866061f0e22e0bb4a82bf9cbbe6c

Observation 4a1f0b84-16fa-41e4-bf57-aae274d446a0 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.781962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:56.914404Z digest=sha256:6e46ce6e754a7fcd806d4df419d5b8d654389a611e92fad7f80898a20b97f81f

Observation 20852e48-d589-4406-8d08-7d1439c0ea86 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.025241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.025241Z digest=sha256:69bf04437403ec60e5a896e119409e7b958e04c80d46db627609a258db9c0a86

Observation 9da3e49d-a12e-44d9-b41c-d7f49378169f · outbound

This paper cites Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT.

Quality Assessment of Python Tests Generated by Large Language Models Evaluating the Code Quality of AI-Assisted Code Generation Tools: An Empirical Study on GitHub Copilot, Amazon CodeWhisperer, and ChatGPT

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:22:57.143131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:22:57.143131Z digest=sha256:04ad71438c8d3c6804653107f401a8a959035db8eff4eb4b3ef532080a9c22b5

Observation 32ddb553-d030-44a7-b15e-9ce263555191 · outbound

This paper cites an unresolved cited work.

Quality Assessment of Python Tests Generated by Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:23:00.629614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:22:57.229784Z digest=sha256:9321324c978c22fd3382b7a7cd4b7e3e5f2f93a470fd0546d6106627f713c60d

Pith citing papers

Observation 039d6874-8fbe-4269-9744-543facd57397 · inbound

Can LLMs be Effective Code Contributors? A Study on Open-source Projects cites this paper.

Can LLMs be Effective Code Contributors? A Study on Open-source Projects Quality Assessment of Python Tests Generated by Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:10.830944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T08:09:33.211692Z digest=sha256:8768ed68d13b4407c43fc4f0a4ef7e4a87d830da4ae1dc6da549e3c8086cca21

Observation a220b4a9-3197-41ba-95b2-36e94fffb31a · inbound

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code cites this paper.

Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code Quality Assessment of Python Tests Generated by Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:26:04.393629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T17:37:51.790000Z digest=sha256:c25a875c97141d2a40d375a11f2063925f1ac144c7f5e4db3b2366c4291b1b7a