Pith. sign in

Paper Citation Record · LEDGER

DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

As of 23 July 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:1903.00161.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1903.00161 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-23T06:31:01.910684+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T18:56:31.772618Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T12:46:14.545000Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 18442e27-26df-414e-9bf3-3837ec838c2d · inbound

How Much Knowledge Can You Pack Into the Parameters of a Language Model? cites this paper.

How Much Knowledge Can You Pack Into the Parameters of a Language Model? DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-15T02:00:28.102026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-15T02:00:28.055865Z digest=sha256:fa6d74343fe6a48b6e879d544002ceef894e206a42fe5c69adc803b20bf0914d

Observation 72ec1ab5-5037-4e3d-99d6-5419ed1bb0de · inbound

Language Models are Few-Shot Learners cites this paper.

Language Models are Few-Shot Learners DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:38.159655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T12:05:38.045330Z digest=sha256:cdb595b6c1448d920f891f636a5c7e78d52344707825f101e987f38b703c8690

Observation 277bd244-2023-4590-ac8c-18656361aded · inbound

Lessons from the Trenches on Reproducible Evaluation of Language Models cites this paper.

Lessons from the Trenches on Reproducible Evaluation of Language Models DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 265

Resolution
verified exact
local_arxiv, observed 2026-05-16T18:44:49.871509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-16T18:44:49.519995Z digest=sha256:cfc396a41e0053cf1ddc59ef2e767d5dba64f69c9d5be5cd8c31ef90ca89953b

Observation 852f9c9f-2f25-4cda-b140-ba5723584466 · inbound

SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model cites this paper.

SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 169

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:30:02.940137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-13T17:30:02.803757Z digest=sha256:5671c5de391dbf7f63d81564480616dd4cbf3df24f786b553523dd73ccf29b1b

Observation bdaee77f-4085-4115-a010-0209cf7afb25 · inbound

Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention cites this paper.

Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T23:46:30.113523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-16T23:46:29.975858Z digest=sha256:73bcd0d52bfbc5e5c90c239714fce69b6df5f80a8619178d853d5ea07981c5d1

Observation 4685d5f5-11a4-4373-8be3-aa889674c79d · inbound

PRIMETIME : Limits of LLMs in Temporal Primitives cites this paper.

PRIMETIME : Limits of LLMs in Temporal Primitives DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-05-22T18:36:58.723207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-22T18:36:48.376877Z digest=sha256:d6c5a95b8cd7b8741c044bc1a3910dadf597f5777d1f9bda52d6687506118c55

Observation d871d4c3-3fbd-4e07-b9fc-f7b840f93e1e · inbound

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review cites this paper.

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 90

Resolution
verified exact
local_arxiv, observed 2026-05-15T02:57:38.372268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-15T02:57:37.873567Z digest=sha256:e41e58d4ffd4b37525d4d9a1dec5b21454b5f64fbdaeb627d9257eda59cb92c6

Observation 742f72b8-7aa5-45e3-9a03-94da8c138675 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-19T11:52:16.451593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:38c863d2366ef7a0d9e48981cd2103b56f467a2b08f189bbd68516239fb500da

Observation 05a06883-8785-4415-8f38-d07b8135575b · inbound

Kimi K2: Open Agentic Intelligence cites this paper.

Kimi K2: Open Agentic Intelligence DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:49:28.112465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T17:49:27.926646Z digest=sha256:0e164ebbd0a2b6c290aa600c4ade6e7d8b1a18d34c7862be61b2e5dbe75a6912

Observation 9aead0b1-ff8f-4e05-92d4-efab434318f8 · inbound

CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit cites this paper.

CreditDecoding: Accelerating Parallel Decoding in Diffusion Large Language Models with Trace Credit DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T09:11:09.969786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-18T09:06:25.049493Z digest=sha256:20e507fc2ccb73d6e3faa0f6491c1cfaf15962f38a31a5d31e444daf27a7d2c6

Observation a1df0756-e16b-4bbd-8222-5383c364e5d1 · inbound

Question-Adaptive Graph Learning for Multi-hop Retrieval Augmented Generation cites this paper.

Question-Adaptive Graph Learning for Multi-hop Retrieval Augmented Generation DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T07:12:27.213571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-18T07:11:48.249136Z digest=sha256:efc24a2da4314233835a00bc3aa61e233cacdf43117662609bd6dba68a2d3f3e

Observation bcbd760a-9eb2-4c63-914b-f6adc2396658 · inbound

Towards Real-World Validity in Generative AI Benchmarks: Understanding and Designing Domain-Centered Evaluations for Journalism Practitioners cites this paper.

Towards Real-World Validity in Generative AI Benchmarks: Understanding and Designing Domain-Centered Evaluations for Journalism Practitioners DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:01:17.034526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-18T10:59:16.139525Z digest=sha256:1786a3fafc29700bfa539a7f446ea9dd0223a9517d6a521b7bb691ba23d21957

Observation e3f9aaf5-65a7-422b-a362-5f48195194ea · inbound

LLaDA2.0: Scaling Up Diffusion Language Models to 100B cites this paper.

LLaDA2.0: Scaling Up Diffusion Language Models to 100B DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-14T18:53:20.984790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-14T18:53:20.911374Z digest=sha256:03996817efad0a81788dab5ed9ee8aeb008e214e2a334b0c2178db07acb6b075

Observation 1e9016f1-0eb7-418e-9bdb-f0be7f70679b · inbound

Learning to Configure Agentic AI Systems cites this paper.

Learning to Configure Agentic AI Systems DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:10:10.382812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-21T13:06:56.207692Z digest=sha256:042f6d1ede5da6859738c18f117b01e717b8625563b4ab65629e8b03cde85d73

Observation 89215d64-6740-4c2d-9025-b2b8e9c8012c · inbound

Learning to Configure Agentic AI Systems cites this paper.

Learning to Configure Agentic AI Systems DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-22T11:31:29.485081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-22T11:29:45.134060Z digest=sha256:8f0d404a8509357d25bf1a2b19cd46627d1df631189b2c9929ec87bb8c388fd3

Observation bf814c69-5073-475e-923f-a8df411c1cca · inbound

SAGE: A Service Agent Graph-guided Evaluation Benchmark cites this paper.

SAGE: A Service Agent Graph-guided Evaluation Benchmark DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:00.828330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T16:41:23.956104Z digest=sha256:ad906bb40a485eb63a9951642f17ee97c51b798de56cc28b31bdd982de68880c

Observation cc9af57f-5a99-4ae7-9632-0c95d610fe98 · inbound

Remask, Don't Replace: Token-to-Mask Refinement in Diffusion Large Language Models cites this paper.

Remask, Don't Replace: Token-to-Mask Refinement in Diffusion Large Language Models DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:29:25.520986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T04:57:48.890310Z digest=sha256:6a7a9c53a2d6ae8404126913beb293203dc1c80eb9e602492837c161372f7856

Observation cdee32ae-9663-4af7-93ab-784ac1c272f0 · inbound

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives cites this paper.

Whose Story Gets Told? Positionality and Bias in LLM Summaries of Life Narratives DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T01:04:50.186727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-10T01:00:41.543394Z digest=sha256:d1c15cd7431c8cb21df983a49d3e1219ef52aec77019a21b45dfd6d08b570b75

Observation 54252ed7-00d8-47f5-afce-a01d43c41e49 · inbound

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks cites this paper.

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:46:05.137085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T00:35:24.397273Z digest=sha256:2048678c02984e9cd8672212182ebc1bb57ab5783a38c1f239a1680b3978b9d0

Observation cd56a7a2-5277-4747-ac92-ed2aa7f1f0fa · inbound

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning cites this paper.

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:26:02.750150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-09T21:53:51.983234Z digest=sha256:0ba387d7a871fbaf4ccdbb5f36dcb48f8b9f37e58ff855ef5f342a0b07caf1c5

Observation 26384852-72c3-4eeb-b7d4-8266c3030fb2 · inbound

Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting cites this paper.

Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:20:42.398356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-08T18:33:26.638240Z digest=sha256:6a8ef0cd3483bddfbd0de9699fd3182f011c7738bc7daf66637f2b7af1c805b2

Observation 67e31863-83c4-4b8f-8af5-53847beaac56 · inbound

MDN: Parallelizing Stepwise Momentum for Delta Linear Attention cites this paper.

MDN: Parallelizing Stepwise Momentum for Delta Linear Attention DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 98

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:41:11.031334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-09T15:27:55.566795Z digest=sha256:2c1ff1100ac61e5327c4507be35130ba05ed010c092c1c9bc699177b2d3df103

Observation 92252072-40c7-404f-860d-3ef2ae8ce598 · inbound

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces cites this paper.

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 67

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.695402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-12T02:57:15.521594Z digest=sha256:9cb4f8b5db2392d5d4ded54592a3652a436d62dc74ce1ce10c2cbafe9253e470

Observation e8a53670-3d65-4f41-976b-b01dd51fc429 · inbound

Reinforced Collaboration in Multi-Agent Flow Networks cites this paper.

Reinforced Collaboration in Multi-Agent Flow Networks DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-14T20:42:57.107466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-14T20:42:48.438057Z digest=sha256:aea4352305da63b07e43d930ee2fcde6f7054950a0d9cba12a46c7fd979b5d64

Observation bb8fe549-6c02-4c77-b717-f97192b0181f · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-20T10:58:14.135248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:682617da8ad428fb7e60315bd62f3b86836c95b5505e0b0d3bc89ef9ec272377

Observation 0328a55f-c599-4b1e-865e-34a1856b1454 · inbound

FedSDR: Federated Self-Distillation with Rectification cites this paper.

FedSDR: Federated Self-Distillation with Rectification DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T12:18:16.300117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-20T12:18:10.362573Z digest=sha256:bcfc2d503ff16cab470954e3e7e12748e760b194437ef2c3f6f3c7ec06d4c42d

Observation 031f3907-89ab-4ed3-a166-a33220117b20 · inbound

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection cites this paper.

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:36:30.306811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-28T09:48:09.688745Z digest=sha256:d2864a2ca43ac523d4f9f5eeeaa7580c3b1f13cbbe54387b82292cbd79822d4d

Observation 73c12433-3a4f-4475-a2fe-626bbca2816a · inbound

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection cites this paper.

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T00:39:16.503684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-07-04T00:30:13.665405Z digest=sha256:4cc20ce06d51434029ea35bbcf29eebf6ffbe5b302393a4760abaada35b5bb0d

Observation aaae6e91-f74c-41fc-9391-86c5613aa0d2 · inbound

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake cites this paper.

LakeQA: An Exploratory QA Benchmark over a Million-Scale Data Lake DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-03T04:47:37.845842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-27T13:41:35.986601Z digest=sha256:1d0c49ce8504c0b4f5511f1ba560fd2bd9f7971f61f12f9090bec1d064cad03c

Observation ae4f12c3-f035-4f87-a136-4dd75027d389 · inbound

Don't Commit Alone: Joint Token Commitment in Diffusion Large Language Models cites this paper.

Don't Commit Alone: Joint Token Commitment in Diffusion Large Language Models DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-11T18:56:31.772618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:56:31.772618Z digest=sha256:5df5b82e636132963f0bde5839674240a41a3715a6e7f55c20a8af1eedea4135

Observation e8948b11-3bb0-434f-b0a9-2c6342e56f83 · inbound

Agentic Data Environments cites this paper.

Agentic Data Environments DROP: A Reading Comprehension Benchmark Requiring Discrete Reasoning Over Paragraphs

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-09T12:46:14.546763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-07-09T12:43:21.070615Z digest=sha256:e5c70c2eda8681d67f8bf6b79f8e8a449e2e7184499e42219ae0496a3ade55cb