Pith. sign in

Paper Citation Record · LEDGER

One Token to Fool LLM-as-a-Judge

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2507.08794.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08794 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T18:53:02.802208Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:15:01.137692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6a3a7621-793b-44ac-98f8-6d3f10c77d70 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge One Token to Fool LLM-as-a-Judge

Reference 218

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:9ef73cfd696a512c1e36e50766a4a3c6b1988c4f676affe745f5b06ae5302115

Observation 0f693e32-c32b-4e35-8cc7-c3c537d19128 · inbound

CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models cites this paper.

CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models One Token to Fool LLM-as-a-Judge

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T18:53:02.802208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:53:02.802208Z digest=sha256:3413daf1d6b95ab1cc1416160b71fc19e7d1ff39b86238fa8cd2e752aa2eced9

Observation 64ea9ca2-05db-47ec-b342-b7b6cf48bc8d · inbound

Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers cites this paper.

Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers One Token to Fool LLM-as-a-Judge

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T07:39:02.616294Z digest=sha256:7dd265e4372ec1bb889d39acddbe021b4f7fa374a03118c1d0f30aa4fefced8f

Observation 0ac07f8c-b302-4334-9815-8be635115515 · inbound

QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs cites this paper.

QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs One Token to Fool LLM-as-a-Judge

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T21:18:03.858801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T21:18:03.858801Z digest=sha256:8bf60161ff1229c7cfc646fa2b42431091664d8cf18752cd86085982e391a541

Observation 178902f2-f9a1-4868-aa97-54674a11cf89 · inbound

Beyond Semantic Manipulation: Token-Space Attacks on Reward Models cites this paper.

Beyond Semantic Manipulation: Token-Space Attacks on Reward Models One Token to Fool LLM-as-a-Judge

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T20:29:31.354743Z digest=sha256:dba99466dcb4c22f3920aa0fde1c3b4672474400661b8a2bbe942e199483acc7

Observation ff8d0b70-579c-460a-9776-745d266ac9dd · inbound

LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection cites this paper.

LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection One Token to Fool LLM-as-a-Judge

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T19:37:19.360378Z digest=sha256:7c33eb7fef53bfc9899f0916d0c1b234196c1732310eb3793fc4c985cff4f23e

Observation 95ea86d6-6133-45e6-9987-8b5466e5ed5a · inbound

Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data cites this paper.

Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data One Token to Fool LLM-as-a-Judge

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T05:32:23.972335Z digest=sha256:77d38fe964408e3fee09016edd906d260c779e91f850d78aff6051da45ef752b

Observation 25ba884c-c147-491b-a6d6-0964f5500f33 · inbound

When AI reviews science: Can we trust the referee? cites this paper.

When AI reviews science: Can we trust the referee? One Token to Fool LLM-as-a-Judge

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T06:19:54.727724Z digest=sha256:9a60d0198289ab705efedccc0f4715ab03de14a9103d2afebb7dd2226cb64ffd

Observation 12f775e4-d555-420d-91f6-8801f592afdd · inbound

Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR cites this paper.

Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR One Token to Fool LLM-as-a-Judge

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T18:52:52.969408Z digest=sha256:b9dcb87d103f4a9c66120c9f605a059116c567c2357046c67746e5f8e0292362

Observation 8d4139fe-d65f-4338-a535-7e6fd2ff8b87 · inbound

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities cites this paper.

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities One Token to Fool LLM-as-a-Judge

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T17:20:39.969447Z digest=sha256:6a5a8fceda2e88b177aa8c5539e12a3f52d705c7dca26e39d7c62138bcb90b3d

Observation d8df954c-e639-43da-8cd7-c6170824e7b9 · inbound

ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization cites this paper.

ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization One Token to Fool LLM-as-a-Judge

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-14T21:16:46.840884Z digest=sha256:899f9b226365c410d7ffd9319d0dd6cef45e66d0a0dabd888dcda42adceb28f6

Observation e9fe35ad-6700-46d4-a903-5eb6b0ad9157 · inbound

ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization cites this paper.

ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization One Token to Fool LLM-as-a-Judge

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-12T02:08:19.458599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T14:33:31.538969Z digest=sha256:c66cc04877bb5baabe11125070b03b01b6e5c1071088fe2a5ab98eadae9e6e91

Observation 2c242ca9-36a2-447c-9e93-c1c0eedc5f54 · inbound

Provably Secure Agent Guardrail cites this paper.

Provably Secure Agent Guardrail One Token to Fool LLM-as-a-Judge

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:53:14.237140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T07:43:17.593344Z digest=sha256:84fc2ae27236bf31372f96ae2f70b1ee8c024ecad6098d5d8b6c5aa8a3972ad1

Observation 9eb52212-92f9-4144-9e99-3215e08fe94d · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation One Token to Fool LLM-as-a-Judge

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T20:56:13.580359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:d802e9d44cdb1c46e4763e4929c41c2252c0664b73cc76e8fe1c4a5086644aa6

Observation 74695256-a15d-4352-8a09-71b667fcdad6 · inbound

Towards Spec Learning: Inference-Time Alignment from Preference Pairs cites this paper.

Towards Spec Learning: Inference-Time Alignment from Preference Pairs One Token to Fool LLM-as-a-Judge

Reference 126

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:39:47.092654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T07:49:36.816100Z digest=sha256:5f987f6e98e96df619cb7b54704cce74dcc028b53e9dcbe9e70e87611d7ca054

Observation 39e102b8-1f61-4ac5-b40f-80bbd850bc37 · inbound

Towards Spec Learning: Inference-Time Alignment from Preference Pairs cites this paper.

Towards Spec Learning: Inference-Time Alignment from Preference Pairs One Token to Fool LLM-as-a-Judge

Reference 126

Resolution
verified exact
local_arxiv, observed 2026-06-30T12:04:39.372749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T10:17:33.176525Z digest=sha256:06cee2e2e353178fe7d1d7f501294d8c01d51f91b0bd4e9d19d23ca69504502a

Observation 4ba8c286-4e8c-454c-9a4d-38776bcd3747 · inbound

From Neural Intent to Cryptographic Authorization: Securing AI-Driven Enterprise Workflows cites this paper.

From Neural Intent to Cryptographic Authorization: Securing AI-Driven Enterprise Workflows One Token to Fool LLM-as-a-Judge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T22:55:45.997271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:55:45.997271Z digest=sha256:c29ac1a2f19051962647271cdeb007a517f9fae114f93ece78a4501632bf994e

Observation ee393540-d4c3-4fec-88af-3568e6d6f7fc · inbound

Codifying the Judge: Scalable Evaluation via Program Distillation cites this paper.

Codifying the Judge: Scalable Evaluation via Program Distillation One Token to Fool LLM-as-a-Judge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T12:49:10.192392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:49:10.192392Z digest=sha256:a2dbe63b0ca2b33b43fd32b9545c93bdf463093e4f091d96c869124c1165e222