Pith. sign in

Paper Citation Record · LEDGER

The Turking Test: Can Language Models Understand Instructions?

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2010.11982.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.11982 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:23:30.529441Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T01:57:29.510754Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d6f3fd2c-1b37-4c28-8422-abf2d4c25962 · inbound

Cross-Task Generalization via Natural Language Crowdsourcing Instructions cites this paper.

Cross-Task Generalization via Natural Language Crowdsourcing Instructions The Turking Test: Can Language Models Understand Instructions?

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:57:29.514479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T01:57:29.380571Z digest=sha256:6d36ee711ce6779573370fb9fa7cfe0d382a17c3a7141cc3edf084c5cd53e36e

Observation feda8011-91fe-42f2-bf49-142a60bc3f3c · inbound

Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? cites this paper.

Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? The Turking Test: Can Language Models Understand Instructions?

Reference 204

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:51:46.883161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-15T09:51:46.701149Z digest=sha256:90e713daf984c0c3b10505b4c798913e277ae6c63229166b62746bd44571a7a4

Observation 64919f7a-36e0-4648-acf0-3d7f50c1df53 · inbound

LIMA: Less Is More for Alignment cites this paper.

LIMA: Less Is More for Alignment The Turking Test: Can Language Models Understand Instructions?

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T11:34:12.937975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-17T11:34:12.854281Z digest=sha256:af95a4ae04bfc9d8ac3426fd7345962226027c21c1b35459c26c1b924a1e56ae

Observation 2d166a2c-eebf-4817-a6c1-c22b73868950 · inbound

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models cites this paper.

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models The Turking Test: Can Language Models Understand Instructions?

Reference 152

Resolution
verified exact
arxiv_id, observed 2026-05-17T14:43:30.189681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-17T14:43:29.496457Z digest=sha256:f9e60dea5a0db95b36a97aec1709cca0e2fb4e34d5b6bd2532018257775a1303

Observation c2f85994-bdc1-4fa4-b79a-7bd2528429b6 · inbound

Multimodal-to-Text Prompt Engineering in Large Language Models Using Feature Embeddings for GNSS Interference Characterization cites this paper.

Multimodal-to-Text Prompt Engineering in Large Language Models Using Feature Embeddings for GNSS Interference Characterization The Turking Test: Can Language Models Understand Instructions?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:30.529441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:30.529441Z digest=sha256:7e9c790ad90198a2a84e3e5d00131e1618338f8dc2d23ca62e4cbe3b35924714

Observation fc7a5f35-5371-4a1f-8b5f-125bf38d7737 · inbound

Visual Textualization for Image Prompted Object Detection cites this paper.

Visual Textualization for Image Prompted Object Detection The Turking Test: Can Language Models Understand Instructions?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:37:03.644571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:37:03.644571Z digest=sha256:1bcec4fe29bcf7c5cbf14e1173c101cb124438f01e74064dbbcc53b95a953b40

Observation f06e3bdf-6aab-4b73-aaf2-fca07a94d057 · inbound

AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration cites this paper.

AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration The Turking Test: Can Language Models Understand Instructions?

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T23:58:42.667440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-16T23:57:19.757902Z digest=sha256:5560b9b4d93c41bafef0f40e3d07401f72f4d81c781f49dc2b94133777ee0f49