Pith. sign in

Paper Citation Record · LEDGER

HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.00215.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.00215 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:42:08.270478Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:57:15.295375Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 42b895fd-3474-4082-9436-f74c37d43d95 · inbound

A Survey on Human-Centric LLMs cites this paper.

A Survey on Human-Centric LLMs HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-12T16:42:08.270478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:42:08.270478Z digest=sha256:a7eee700886f2cd709feefa34efb7170bd19bb6523599d3b03de3fc1bbcdb743

Observation adcbdedf-7b55-4817-b71f-dfd011f56d36 · inbound

Unseen Horizons: Unveiling the Real Capability of LLM Code Generation Beyond the Familiar cites this paper.

Unseen Horizons: Unveiling the Real Capability of LLM Code Generation Beyond the Familiar HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T18:16:26.812431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:16:26.812431Z digest=sha256:affaea6fce9180737df624c7340b0d93229c193869a1b22e7adb2d8daf5872d7

Observation c8b9e154-3692-4a8c-9490-6c60343e7315 · inbound

AuPair: Golden Example Pairs for Code Repair cites this paper.

AuPair: Golden Example Pairs for Code Repair HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T05:45:13.330877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T05:45:13.330877Z digest=sha256:8e50da6265a1d8465024a1da83bd93fb8b8df1bb56d7b671d2d35e90190402d9

Observation 52f6669d-fd95-40f6-90f1-ae156aafe0ea · inbound

Assessing the Impact of Requirement Ambiguity on LLM-based Function-Level Code Generation cites this paper.

Assessing the Impact of Requirement Ambiguity on LLM-based Function-Level Code Generation HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:41:21.166055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T21:15:51.570343Z digest=sha256:58261d0836706261c9897248a5f0292b9702d9cdd6a1ec200e8b136cf86aaf1c

Observation f54e13d0-9d9b-463e-8069-07b8c929561a · inbound

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback cites this paper.

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:23:10.625954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T09:22:06.285118Z digest=sha256:92d27f56360d0669d29246d4d0d74de90383fdab28f4ab7151bc33cb2dda4ee9

Observation 4afe3cb0-1d16-4036-8d69-f44e9fb255f8 · inbound

The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code cites this paper.

The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:57:15.296775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-02T17:50:07.458116Z digest=sha256:fc9c78e15aa6b14253e75de32ac4b954c802d8bc48cc2e38d50ed635c9a4a74a