Pith. sign in

Paper Citation Record · LEDGER

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models

As of 22 August 2026, this Paper Citation Record lists 10 of 10 outbound references and 1 inbound Pith citation observation for arXiv:2507.15868.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15868 v1

Coverage vector

measured 10 of 10 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:26:58.398605Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T17:43:48.456679Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T17:49:23.561376Z

Reference resolution

10 of 10 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a497670-7b41-4cbb-b7e6-25af520195d5 · outbound

This paper cites LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:57.798530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:57.798530Z digest=sha256:3a351a7ed304c4347682532e62c8e0ca91788c0d1fab43504c24a7b1b1aa832b

Observation 545bc79d-90d7-4a0c-a540-2cf8b4bf5d9d · outbound

This paper cites Halstead.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Halstead

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:59.132730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:26:57.884653Z digest=sha256:811a375a8c70cd01870053dfd30f8ec76cbe21602fff84a0495cb08ad80696fc

Observation 99e5430a-a6ad-4153-b26f-3c63df448ba0 · outbound

This paper cites Adversarial ex- amples for evaluating reading comprehen- sion systems.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Adversarial ex- amples for evaluating reading comprehen- sion systems

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:59.000290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:26:57.959899Z digest=sha256:052938cd4fbc59001e24ba26ba05060a4b7399fcbd21b11f983df23c771444ce

Observation e45d585c-6a1d-4fc7-ab1c-e4c385131356 · outbound

This paper cites On faith- fulness and factuality in abstractive sum- marization.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models On faith- fulness and factuality in abstractive sum- marization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:58.905222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:26:58.029361Z digest=sha256:ca83abc1c60566bcb74a4be95f7d83ef0b6a1f1eed44958c7e55abd98ff63c6d

Observation 6aad6d91-519d-4a76-ad5f-2a5e53c6c458 · outbound

This paper cites Assessing Hidden Risks of LLMs: An Empirical Study on Robustness, Consistency, and Credibility.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Assessing Hidden Risks of LLMs: An Empirical Study on Robustness, Consistency, and Credibility

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.094112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.094112Z digest=sha256:7e98b1f1d6f566f65d78e12306e9ab827d4ae9729dd65b0e2b1c768f9656fe12

Observation 40cacaab-bdfc-4206-80c1-3fe77daf5e67 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.175019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.175019Z digest=sha256:e14be6ab482ac757314413a9f8f55463bb7cf9a7625bbef68937fc0df5b3f0b5

Observation 09b41dde-eeb1-464e-b564-beb7cf2a7739 · outbound

This paper cites Univer- sal adversarial triggers for attacking and an- alyzing NLP.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Univer- sal adversarial triggers for attacking and an- alyzing NLP

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:58.771370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:26:58.222319Z digest=sha256:9ed5c55619816e29d986515a984e9181fe63d92042a4135ffcee3ef946009f8a

Observation 99e8532f-c972-49c5-884b-903363f99cc0 · outbound

This paper cites Don’t take the easy way out: Ensemble based methods for avoid- ing known dataset biases.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Don’t take the easy way out: Ensemble based methods for avoid- ing known dataset biases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:26:58.623531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:26:58.300798Z digest=sha256:c1f4c919633eaaba9c42de16fb1f67df94fbfed665a7231c0f4cd141e0b6e0c9

Observation 3c6a8bfe-7fae-4a1d-9d39-17efd68b6a66 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.346488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.346488Z digest=sha256:9bffdee5bc6ea5c7467d4fbdabac2ca6685073e3aec9920ec37d71236ce6c76d

Observation db0697e7-fef1-419d-8dab-d1b460a130e7 · outbound

This paper cites Holistic Evaluation of Language Models.

Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models Holistic Evaluation of Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:58.398605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:58.398605Z digest=sha256:b02ac322d4942f912b6372611127cc2850a637bce23775338a6da28e84b56075

Pith citing papers

Observation 97a4f4e8-8bda-4573-83b9-0eee35b170c4 · inbound

Learning Perturbations to Extrapolate Your LLM cites this paper.

Learning Perturbations to Extrapolate Your LLM Small Edits, Big Consequences: Telling Good from Bad Robustness in Large Language Models

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:49:23.563898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-14T17:43:48.456679Z digest=sha256:345b96485abe5a31001a8fb771829ccadb0de4e198b426fde98828a1588739a3