Pith. sign in

Paper Citation Record · LEDGER

PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.19740.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19740 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T13:40:46.709807Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:06:35.205226Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9011def-dae2-4f1d-87ff-a979333e27b6 · inbound

LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient cites this paper.

LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T18:09:01.639375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T18:09:01.639375Z digest=sha256:ec34efe86ef2ab866e89f476fdfb87fdfd235287fa2295f110b79088ab81f774

Observation 9a3bb878-238b-4286-a246-8037bd818d26 · inbound

Governing AI Beyond the Pretraining Frontier cites this paper.

Governing AI Beyond the Pretraining Frontier PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:46.709807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:46.709807Z digest=sha256:13dfb76b4f7577b380f7196a6db8ccbac2aa2de0a6f3abe862b9e42f944af78f

Observation 53df4759-7c2d-4fb4-9e65-93c92e196499 · inbound

AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data cites this paper.

AutoEvoEval: An Automated Framework for Evolving Close-Ended LLM Evaluation Data PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:01.719577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:01.719577Z digest=sha256:2db9667b3d4194ce4d532be16741aa09406db7b9671a2067d1fabb6c71448cb5

Observation 53967675-f847-4a91-9882-90e335104bdf · inbound

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules cites this paper.

DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rules PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:31:25.892228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T01:00:13.017290Z digest=sha256:06da7a109b1c5b9fb67b27ac5d9c55cc8138f85ee6f8ad27b9c5d987b29f3a1d

Observation c399ec89-b0ad-42a5-a561-b12ffff554c6 · inbound

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks cites this paper.

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:06:35.206717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T09:27:30.923556Z digest=sha256:bb22bfd7766de1600ac7038328ad0938f5c574f95be453ab2a07de8793e196c8