Pith. sign in

Paper Citation Record · LEDGER

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models

As of 7 August 2026, this Paper Citation Record lists 10 of 10 outbound references and 1 inbound Pith citation observation for arXiv:2507.15868.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15868 v1

Coverage vector

measured 10 of 10 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:26:58.398605Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T17:43:48.456679Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T17:49:23.561376Z

Reference resolution

10 of 10 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a497670-7b41-4cbb-b7e6-25af520195d5 · outbound

This paper cites LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:57.798530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:57.798530Z digest=sha256:8d0ed0a256db112347e546e2415c2071924f7093bf188c8384dd82e9f874a334

Observation 545bc79d-90d7-4a0c-a540-2cf8b4bf5d9d · outbound

This paper cites Halstead.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Halstead

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:59.132730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:57.884653Z digest=sha256:8b35c72d4d1e7253b89416533080d319557a864bdf6de7e3d25ebd320144355b

Observation 99e5430a-a6ad-4153-b26f-3c63df448ba0 · outbound

This paper cites Adversarial ex- amples for evaluating reading comprehen- sion systems.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Adversarial ex- amples for evaluating reading comprehen- sion systems

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:59.000290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:57.959899Z digest=sha256:8c1669c5dfcd197f15c2c74870c12e506f02a2a0326439f2d9b6418c71dedb86

Observation e45d585c-6a1d-4fc7-ab1c-e4c385131356 · outbound

This paper cites On faith- fulness and factuality in abstractive sum- marization.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models On faith- fulness and factuality in abstractive sum- marization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:58.905222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:58.029361Z digest=sha256:29b4b39620f27bed126b701de9d9b1c4286e1002957f89a27baddfe90dba61fa

Observation 6aad6d91-519d-4a76-ad5f-2a5e53c6c458 · outbound

This paper cites Assessing Hidden Risks of LLMs: An Empirical Study on Robustness, Consistency, and Credibility.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Assessing Hidden Risks of LLMs: An Empirical Study on Robustness, Consistency, and Credibility

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.094112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.094112Z digest=sha256:3972ad2cfcd9bff6bad8f08ec6d8c2b8c1765fee6ab1398cc70307b5e66bcaf3

Observation 40cacaab-bdfc-4206-80c1-3fe77daf5e67 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.175019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.175019Z digest=sha256:280dabf0ce2cdae7751f723476506379e016e485f7eb2af825a21c01a1ccdac2

Observation 09b41dde-eeb1-464e-b564-beb7cf2a7739 · outbound

This paper cites Univer- sal adversarial triggers for attacking and an- alyzing NLP.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Univer- sal adversarial triggers for attacking and an- alyzing NLP

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:58.771370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:58.222319Z digest=sha256:ae9866cb60a35c3a5891dcc0a44ed07d3d204a2bfc624a8af9ff37254dd74bd9

Observation 99e8532f-c972-49c5-884b-903363f99cc0 · outbound

This paper cites Don’t take the easy way out: Ensemble based methods for avoid- ing known dataset biases.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Don’t take the easy way out: Ensemble based methods for avoid- ing known dataset biases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:58.623531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:26:58.300798Z digest=sha256:fcc5a0728afa4049b5e1ac12e7bd9abfa03ebbc0681d7fb9048fb91da3c8842e

Observation 3c6a8bfe-7fae-4a1d-9d39-17efd68b6a66 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.346488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.346488Z digest=sha256:e69252437d44cfa24e7c70a140675b2698ec8b594712c7b2e4f7f701fd79349b

Observation db0697e7-fef1-419d-8dab-d1b460a130e7 · outbound

This paper cites Holistic Evaluation of Language Models.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Holistic Evaluation of Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.398605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.398605Z digest=sha256:3e021ba528279f7b2c44066a0d4e9f4dafe4bd83f01245ed93121b6259bc28c4

Pith citing papers

Observation 97a4f4e8-8bda-4573-83b9-0eee35b170c4 · inbound

Learning Perturbations to Extrapolate Your LLM cites this paper.

Learning Perturbations to Extrapolate Your LLM Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:49:23.563898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T17:43:48.456679Z digest=sha256:0c760cc4edab07fa08b53bd534849ca063d33e263a059ecda6116eb8d371437b