Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:20:57.144292Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2412.15275.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:20:57.144292Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 004b0fea-5fae-4b7c-8cc5-253764f1ff5d · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting Inconsistency in Conference Peer Review: Revisiting the 2014 NeurIPS Experiment
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48f38db0-4242-4e24-9c16-acffe74ecd80 · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting The Llama 3 Herd of Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaafae0c-e69d-4cd5-b876-edf9152585e4 · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6502c70-a5a6-4672-ab23-cb511016746f · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a64259a-ae68-4911-8261-1b3db2e1d91e · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting WizardLM: Empowering large pre-trained language models to follow complex instructions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 022b811b-e11a-473e-8950-602c5a35e4ea · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72826b2b-ee2c-497e-b30e-0b5a3104162c · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting Representation Engineering: A Top-Down Approach to AI Transparency
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18df5ae2-02d8-4966-9417-f705327413f5 · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting Format Specification We provide a format to follow
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 320ec835-f5b7-490a-b147-bdca7077e3d4 · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting The Hewlett Foundation: Automated Essay Scoring
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 191b4587-f3e8-472f-ba19-da6cbf7f9f05 · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting Gemini: A Family of Highly Capable Multimodal Models
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fad9620-c98b-4633-ae36-72b739c0c8ac · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting Revisiting Jailbreaking for Large Language Models: A Representation Engineering Perspective
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ec5a374-ae7b-4ce0-80a5-cf6e1507a907 · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting org/blog/2023-03-30-vicuna,
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 70898b16-8d22-4185-90bb-b48d00d0463a · outbound
Fooling LLM graders into giving better grades through neural activity guided adversarial prompting RLHF Workflow: From Reward Modeling to Online RLHF
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.