Pith. sign in

Paper Citation Record · LEDGER

All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2107.00061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2107.00061 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T21:10:56.960278Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T22:40:23.517409Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6fbd76d3-6fbf-4d17-82da-ae69a47c552e · inbound

Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages cites this paper.

Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T21:10:56.960278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:10:56.960278Z digest=sha256:d5b4f8a8882bddd42e3922731d4c959c9b908112e5d0b3dcd952996fb29e3ee2

Observation 7490b38a-af62-4ccf-a9d7-103b248659a6 · inbound

Is Your LLM-Based Multi-Agent a Reliable Real-World Planner? Exploring Fraud Detection in Travel Planning cites this paper.

Is Your LLM-Based Multi-Agent a Reliable Real-World Planner? Exploring Fraud Detection in Travel Planning All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:57.959163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:57.959163Z digest=sha256:43ab54050afda67af827ae3c5b3f7b00870bf3f2f25c75f1e629c948d5e367f6

Observation 29a15694-57eb-46dc-b94e-a18d5928a149 · inbound

Towards Efficient and Effective Alignment of Large Language Models cites this paper.

Towards Efficient and Effective Alignment of Large Language Models All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:36.134233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:36.134233Z digest=sha256:93aaaf579ac8c3a0b6920db17522ffb0f0647f1fc2438c9127fbd9d8b48006ea

Observation dcfb530f-5714-443e-a72e-85a4f09eeebb · inbound

Psychology-Driven Enhancement of Humour Translation cites this paper.

Psychology-Driven Enhancement of Humour Translation All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:30.342101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:04:30.342101Z digest=sha256:a22b0b811840c137f3ac5e6f74987818efc713521e11bd9c2c6b3d7177d252fa

Observation 920fb70b-99d7-44d0-aba9-f026b01f728b · inbound

LLM Encoder vs. Decoder: Robust Detection of Chinese AI-Generated Text with LoRA cites this paper.

LLM Encoder vs. Decoder: Robust Detection of Chinese AI-Generated Text with LoRA All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T13:22:08.112792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:22:08.112792Z digest=sha256:81922402eb8875451e0320b473e8e3c989c8ce3a19dbbbf5f5040d45adf8da2d

Observation 8c437b23-020b-4828-80ad-21f13f2a5799 · inbound

Towards Trustworthy AI: Characterizing User-Reported Risks across LLMs "In the Wild" cites this paper.

Towards Trustworthy AI: Characterizing User-Reported Risks across LLMs "In the Wild" All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:03:35.978854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:03:35.978854Z digest=sha256:c163ff9af7ef5278dd8526129ec06a2c223eeb15f7574e73dc874f3b65a0a75b

Observation 87547f1c-5e92-4d35-959a-1c915fc18786 · inbound

Detecting LLM-Assisted Academic Dishonesty using Keystroke Dynamics cites this paper.

Detecting LLM-Assisted Academic Dishonesty using Keystroke Dynamics All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:40:23.520528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T22:39:54.384363Z digest=sha256:bcb5fcecfa0aa165f8499261f510e65e5b57d4391cd72b752f0591ea00d9485a

Observation 18d0f027-7e4e-4119-b3e4-3b443da255cf · inbound

Network Effects and Agreement Drift in LLM Debates cites this paper.

Network Effects and Agreement Drift in LLM Debates All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:46:08.002724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:50:20.833398Z digest=sha256:d47717bd86003c802e067de3ac67589e65c7d323720a7ba5e58dc0223b113db2

Observation f065e130-ba71-43de-8fa9-6c8a0cfac72e · inbound

Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild cites this paper.

Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:27:48.159299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T11:26:38.634540Z digest=sha256:0c695fd24c98b918eae394fb393f1f849e64002cc244e03e8cb03dd3d1c4255b

Observation 029053f4-a196-4af4-807d-4e3d9f5c1021 · inbound

The unintended consequences of large language models as a labor-augmenting technology in science cites this paper.

The unintended consequences of large language models as a labor-augmenting technology in science All That's 'Human' Is Not Gold: Evaluating Human Evaluation of Generated Text

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-01T18:10:24.264365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:10:24.264365Z digest=sha256:8bf170a5a486b98f1bdba06e54566b7925c09aeebd0d8dfa094d10b15bbf9d84