Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2507.08794.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T18:53:02.802208Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T12:15:01.137692Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 6a3a7621-793b-44ac-98f8-6d3f10c77d70 · inbound
Reference 218
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0f693e32-c32b-4e35-8cc7-c3c537d19128 · inbound
CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models One Token to Fool LLM-as-a-Judge
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ea9ca2-05db-47ec-b342-b7b6cf48bc8d · inbound
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers One Token to Fool LLM-as-a-Judge
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0ac07f8c-b302-4334-9815-8be635115515 · inbound
QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs One Token to Fool LLM-as-a-Judge
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 178902f2-f9a1-4868-aa97-54674a11cf89 · inbound
Beyond Semantic Manipulation: Token-Space Attacks on Reward Models One Token to Fool LLM-as-a-Judge
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ff8d0b70-579c-460a-9776-745d266ac9dd · inbound
LLM-as-Judge for Semantic Judging of Powerline Segmentation in UAV Inspection One Token to Fool LLM-as-a-Judge
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 95ea86d6-6133-45e6-9987-8b5466e5ed5a · inbound
Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data One Token to Fool LLM-as-a-Judge
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 25ba884c-c147-491b-a6d6-0964f5500f33 · inbound
When AI reviews science: Can we trust the referee? One Token to Fool LLM-as-a-Judge
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 12f775e4-d555-420d-91f6-8801f592afdd · inbound
Delay, Plateau, or Collapse: Evaluating the Impact of Systematic Verification Error on RLVR One Token to Fool LLM-as-a-Judge
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8d4139fe-d65f-4338-a535-7e6fd2ff8b87 · inbound
Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities One Token to Fool LLM-as-a-Judge
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d8df954c-e639-43da-8cd7-c6170824e7b9 · inbound
ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization One Token to Fool LLM-as-a-Judge
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e9fe35ad-6700-46d4-a903-5eb6b0ad9157 · inbound
ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization One Token to Fool LLM-as-a-Judge
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2c242ca9-36a2-447c-9e93-c1c0eedc5f54 · inbound
Provably Secure Agent Guardrail One Token to Fool LLM-as-a-Judge
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9eb52212-92f9-4144-9e99-3215e08fe94d · inbound
Trust Region On-Policy Distillation One Token to Fool LLM-as-a-Judge
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 74695256-a15d-4352-8a09-71b667fcdad6 · inbound
Towards Spec Learning: Inference-Time Alignment from Preference Pairs One Token to Fool LLM-as-a-Judge
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 39e102b8-1f61-4ac5-b40f-80bbd850bc37 · inbound
Towards Spec Learning: Inference-Time Alignment from Preference Pairs One Token to Fool LLM-as-a-Judge
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ba8c286-4e8c-454c-9a4d-38776bcd3747 · inbound
From Neural Intent to Cryptographic Authorization: Securing AI-Driven Enterprise Workflows One Token to Fool LLM-as-a-Judge
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee393540-d4c3-4fec-88af-3568e6d6f7fc · inbound
Codifying the Judge: Scalable Evaluation via Program Distillation One Token to Fool LLM-as-a-Judge
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.