Pith. sign in

Paper Citation Record · LEDGER

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis

As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2508.03396.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.03396 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T04:35:32.587153Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6efec82b-67aa-4172-a0d4-c65cb0d0d945 · outbound

This paper cites S.; Kazi, S.; Sourabh, V.; et al.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis S.; Kazi, S.; Sourabh, V.; et al

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:36.239512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.014179Z digest=sha256:7a0fba297095c8337615345c8752ed7a12ea3976e14311211925727f14f9cad5

Observation 53e67fae-f0c7-4a85-9e3f-8cb8ddf42f6b · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:36.031185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.028774Z digest=sha256:ce6d83a6deb2b16d75bcac615d5c73dccecff42beba42118de192fa84dad2266

Observation f732523d-7f1e-43f9-9c1b-b92c835bcf0a · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:35.903118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.042241Z digest=sha256:2b6ad1f504344017399043667e3ad31284cccff106ae35620dd97d362a6d8d73

Observation 56a9be93-3491-44ef-ba28-da5c3a95d859 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.057805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.057805Z digest=sha256:e29e993e2232424811a222de97d5d5d36cd485e4bfa8d36ad357ab17ff184141

Observation dfd12a22-b3f9-4d70-9b11-dfb5301856d3 · outbound

This paper cites UloRL:An Ultra-Long Output Reinforcement Learning Approach for Advancing Large Language Models' Reasoning Abilities.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis UloRL:An Ultra-Long Output Reinforcement Learning Approach for Advancing Large Language Models' Reasoning Abilities

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.077918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.077918Z digest=sha256:e8985929b1d2b82dc7da0ca29a4363ddede565fa773633bd3a7daa157d1cfd83

Observation 25dbc973-0abc-41cc-a0ef-2b700ecb86aa · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:35.731083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.091051Z digest=sha256:e935d9781b2f8f1d704dcaab7c0a55e75b03b4fa0b18dfe7d2acb76e32720870

Observation 8a565c9e-096e-401a-9847-6e190f6c7802 · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:35.560043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.104790Z digest=sha256:2d17dbbec8c1c0b2800a2236ed58fa90dd4427fd8c06b14ce3fe191bd34251a7

Observation 7ec34be1-9257-49ef-9de5-b01380e0090b · outbound

This paper cites L.; Liu, Y.; Shang, N.; Sun, Y.; Zhu, Y.; Yang, F.; and Yang, M.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis L.; Liu, Y.; Shang, N.; Sun, Y.; Zhu, Y.; Yang, F.; and Yang, M

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:35.370968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.119363Z digest=sha256:a43ac2b674792870451c3d41e9df1b1bfbb0ff89574279d9339878d77f24e1f9

Observation 3c1e1432-07d5-4af5-82dd-b8b1470b9a3d · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:35.157693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.134048Z digest=sha256:db632e6c61ff07566deca8561463962ab69fd6c5427b1399c18a346d9ca7a47e

Observation 9b9db833-d11d-4a2f-850c-9221e873f626 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Measuring Mathematical Problem Solving With the MATH Dataset

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.149105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.149105Z digest=sha256:c906ea384e946f6f2b4e8a5082fe5a1dc604371426a6e9e6bfac3379b3672061

Observation de8a7881-3eec-4a4d-8712-61e6474933e2 · outbound

This paper cites S.; Yu, A.; Song, X.; and Zhou, D.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis S.; Yu, A.; Song, X.; and Zhou, D

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:34.828222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.164078Z digest=sha256:ecc4f78283c0bcd5d8e2386fb94ab944e84c2ed92263a55291cb43af5cf76d90

Observation 915c15dc-8988-4842-9e87-dd6f040bed42 · outbound

This paper cites Prover-Verifier Games improve legibility of LLM outputs.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Prover-Verifier Games improve legibility of LLM outputs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.176661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.176661Z digest=sha256:d50ba12d73cbc6cfe280bb5890824a326b4c7d885310d64bf07774d288314488

Observation ee58e9c9-b9d9-4112-8a18-fd3b5edd409b · outbound

This paper cites Q.; Shen, Z.; et al.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Q.; Shen, Z.; et al

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:34.577326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.191255Z digest=sha256:26fcfe2cd149711ffcd12e2f3bbd2296c63cc338935fcb9f7a91b1cb23dd4a96

Observation 3e0f0473-5d9b-454c-8cd9-d5542d5d5485 · outbound

This paper cites MathDebugger: Detecting and Diagnosing Errors in Synthetic Mathematical Data.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis MathDebugger: Detecting and Diagnosing Errors in Synthetic Mathematical Data

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.214664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.214664Z digest=sha256:d3c15e8392827f32f8c2249c0ca0fe6ffc2a319a5ecb45ed0c90184a776b6e64

Observation 8f0452db-b52a-4391-ae76-d4125b656030 · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:34.296698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.232841Z digest=sha256:7c880f259bcdd1d0fa73346fa76f4546ccfd5d69b99f4baee5be8b7a0a36d65a

Observation 17b8ece2-d40e-4eeb-b0d3-f335fa4b077d · outbound

This paper cites Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergy.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.254766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.254766Z digest=sha256:e97a2405f7535ce522e7142f40783d54f401a214708c5c756b459c2d0a5cad84

Observation 2c50cb34-9d42-4ce3-9899-47a61997746e · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:33.972176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.345470Z digest=sha256:32d6804272af3a0a60fd1109cc00c82b357acb2820e8b861427219b739e3660c

Observation fb554cd4-817e-49dd-b1a0-a525861ecc07 · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:33.795569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.453897Z digest=sha256:39d9f75969f51daabb8e3a1b3784877b500213f7aec975a15327ead42e60eefa

Observation c26cf467-c98f-434c-9c75-0eb38ec3faf9 · outbound

This paper cites D.; Ermon, S.; and Finn, C.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis D.; Ermon, S.; and Finn, C

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:33.633698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.523214Z digest=sha256:7dfbfe53fe03efa99b338170808cbd5362168609342ff4effac966d3b7275a92

Observation 1800bddd-1402-434d-a15b-9317b518e41e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Proximal Policy Optimization Algorithms

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.595403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.595403Z digest=sha256:be2a108b86263303124caa72092e2be097bff4099fffc8a40c4dca39eedf6162

Observation c1dd8568-a275-4aa0-8d9b-9ba935fc77f1 · outbound

This paper cites On a Formal Model of Safe and Scalable Self-driving Cars.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis On a Formal Model of Safe and Scalable Self-driving Cars

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.690384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.690384Z digest=sha256:ff46d67ec507e52836c51e44f05afc0ea60e7a3297ed4562ea558d71b19aa6d1

Observation 8409a533-71ac-4edd-bc3d-3e125a4e41ea · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:31.762696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:31.762696Z digest=sha256:ad404d3bcb57dbff718658f7826f1ac4dae8f9d2e9e4ce10c8211a03d6cc06ce

Observation 070a0122-2a06-4a99-909a-3472a0bd88f3 · outbound

This paper cites P.; and Mak, T.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis P.; and Mak, T

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:33.464339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.863051Z digest=sha256:7eaa6feb6232730ae9f3491d2753b6267f4f57cafd6025d095f40a1fae388579

Observation 30ce62c6-35bb-45b8-ae91-7e12a979ceed · outbound

This paper cites an unresolved cited work.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T04:35:33.323862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:31.977617Z digest=sha256:6e891d825a4f5b9714a11cc3e82c890bad03d83c9ed5a7cf616b97722dd7daea

Observation 3413d6a4-733c-4b9c-a08e-e35f5047d21c · outbound

This paper cites Large Language Models for Education: A Survey and Outlook.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Large Language Models for Education: A Survey and Outlook

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:32.045533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:32.045533Z digest=sha256:fda3931affc9b73839305c8c47fb000a84ca788a9cdb2a36dba08f17cdebfd89

Observation 883de42d-e437-4fd9-b43f-1b2727c879b5 · outbound

This paper cites V.; Zhou, D.; et al.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis V.; Zhou, D.; et al

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T04:35:33.181720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T04:35:32.119965Z digest=sha256:a6bd7e2e447c123ecb82b6fb44feb991d5a22eb0256c1d38df0ef92e0d57f5e2

Observation 05b28a88-1c5d-4b92-8407-93b2008ea8d3 · outbound

This paper cites Unveiling the Safety of GPT-4o: An Empirical Study using Jailbreak Attacks.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Unveiling the Safety of GPT-4o: An Empirical Study using Jailbreak Attacks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:32.226777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:32.226777Z digest=sha256:3deb2dd4da344a279452161633b7084876ecb4cb71b6e704df5178eb9f924722

Observation b29c6c51-9e65-48f3-8de4-4a48c5ab1f92 · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:32.297735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:32.297735Z digest=sha256:b94601eafcda5238586dadb2ee870259dbfb70f11d40f52ee73a5d741e65448c

Observation c57377c5-daf4-4b04-b091-0e812290174e · outbound

This paper cites HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:32.417153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:32.417153Z digest=sha256:9bc95d79dae60bdc514cf13087ceddd1322ce7bd2f7311bc45e3a312b3ea7340

Observation 71eeedfb-b77b-43d5-8d65-75837760cc9d · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis , " * write output.state after.block = add.period write newline

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:32.488758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:32.488758Z digest=sha256:3363d29a71eac3b0f631583a4e8191a2587ff11708dd536ecba1e1430428c055

Observation 80bbb304-1516-4d8e-8fd1-0370ae92a777 · outbound

This paper cites write newline.

Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis write newline

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T04:35:32.587153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:35:32.587153Z digest=sha256:532229ec0ef68f128c5394ebd03b0e774b7076488ba047a3b85039db5b535036

Pith citing papers

No inbound Pith citation observations are available.