Pith. sign in

Paper Citation Record · LEDGER

Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2402.14016.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.14016 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:52.421255Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T17:35:43.987419Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04588e05-18b4-49ea-860c-c307ab898808 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 119

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:43.990648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:89558714b884c70d8e7b2abf5aee7de7dfe714d87e7919495244699584f0d525

Observation 304c9cd5-a2d1-453d-9854-a2a083b98a4e · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:34.639350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:538ff4788e68ab10b85bad56fee90dad755410c4161efa628d9156f08e2aa3bc

Observation e39f4052-00bc-4d4b-9d89-f8c258d4d13e · inbound

Towards Understanding the Robustness of LLM-based Evaluations under Perturbations cites this paper.

Towards Understanding the Robustness of LLM-based Evaluations under Perturbations Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T17:09:54.139224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:09:54.139224Z digest=sha256:aa4133bae3c155620187a4e6b0677ce37fc7502f8ab9eef0af6c7c88287ab705

Observation a5657ff9-64fe-44bd-9873-1096621b70c6 · inbound

Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-box Neural Ranking Models cites this paper.

Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-box Neural Ranking Models Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T04:32:36.186677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:32:36.186677Z digest=sha256:58c970ab18f7d4d4c02b443e9b885ef5c3759fe9d04162473f12b598084de9e0

Observation f57be3d8-6ebd-4101-85a2-cbe6faa4160e · inbound

Chatperone: An LLM-Based Negotiable Scaffolding System for Mediating Adolescent Mobile Interactions cites this paper.

Chatperone: An LLM-Based Negotiable Scaffolding System for Mediating Adolescent Mobile Interactions Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:52.421255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:28:52.421255Z digest=sha256:87a5b005d10601dad7fc4e96fba84bb259bb2d503a152558089afe31bebb48ee

Observation df39e853-63f1-4ca5-91a7-e98798b257da · inbound

The Leaderboard Illusion cites this paper.

The Leaderboard Illusion Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-16T05:22:55.621429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:22:55.621429Z digest=sha256:8e7e0f698ee9705ca60c41b1168f886e1b5807be8e5efe5d84d867bb8b9668ed

Observation 00788e6f-822a-407f-92e2-abf6df22ed5b · inbound

One Token to Fool LLM-as-a-Judge cites this paper.

One Token to Fool LLM-as-a-Judge Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:38.612180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:15:38.612180Z digest=sha256:27f80b88367ba98b5162a5ad675bd9ed2994f5466b696fd7ba68853e9bfc7eb7

Observation 251f3ac3-81a6-405d-b40c-8702098bca6a · inbound

TripTailor: A Real-World Benchmark for Personalized Travel Planning cites this paper.

TripTailor: A Real-World Benchmark for Personalized Travel Planning Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T05:40:31.814090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:40:31.814090Z digest=sha256:7400e3eae3b15298d8c1683c18de562d0816d2ed241c0a834204bdfb1ddd5842

Observation ac8e5f32-794d-40f0-ab27-6d0b0dd180b1 · inbound

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization cites this paper.

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.605293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T12:53:45.767341Z digest=sha256:615f75e913fbcb803ad47d7e48844944709c356d5cadfe0f173cc9ac5569984a

Observation 73696ae7-9264-4465-9555-69c1a26a9c9f · inbound

When AI reviews science: Can we trust the referee? cites this paper.

When AI reviews science: Can we trust the referee? Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 111

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:16:11.606972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T06:19:54.727724Z digest=sha256:78f110f94a4a1f44a6fb4fe9c12f8164076ceb79cadc41d833449fb88134d613

Observation a4435910-bfd7-4897-827d-2203faf2ba23 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:20.567329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:20.567329Z digest=sha256:4f0c7eb300bfecb9d6e0a65c3ca945e92e9f1e7c4035704497c394a63b8e4074

Observation 04ce2feb-6030-4475-83fc-3616d40e9a35 · inbound

V-FiLLM: Verified Financial LLM Reasoning Benchmark cites this paper.

V-FiLLM: Verified Financial LLM Reasoning Benchmark Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T11:27:14.659813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T11:27:14.659813Z digest=sha256:77f175bb4a8540788e563f3f282501cf7f86c6d43a6ed05e97a99bb95aa60a9e