Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T11:30:35.285902Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2606.04075.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T11:30:35.285902Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9efdbd50-f087-4816-a5da-91dea8936e6c · outbound
Large Language Models Hack Rewards, and Society Concrete Problems in AI Safety
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54b98638-9d67-420f-afb9-6aeb6d0faf41 · outbound
Large Language Models Hack Rewards, and Society Advances in Neural Information Processing Systems , volume=
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c5a8d67-acc0-40ef-a2ec-93e5c1e1ac91 · outbound
Large Language Models Hack Rewards, and Society International Conference on Learning Representations , year=
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 584b8cfb-ef6e-4f39-a031-7560d8c3f61d · outbound
Large Language Models Hack Rewards, and Society Specification gaming: the flip side of
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8a451c3-6c4c-4b86-8401-3be7115aae5e · outbound
Large Language Models Hack Rewards, and Society Advances in Neural Information Processing Systems , volume=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffb2b9c2-dc12-4877-85fd-c94abdde8eb9 · outbound
Large Language Models Hack Rewards, and Society Advances in Neural Information Processing Systems , volume=
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a22189c9-f4c2-4e04-abb8-56a805d470f9 · outbound
Large Language Models Hack Rewards, and Society Advances in Neural Information Processing Systems , volume=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c85e5c66-30a2-43f7-abe0-e5797eb5a412 · outbound
Large Language Models Hack Rewards, and Society Constitutional
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5afb8ea0-e3b8-48da-8dd8-a819b67635f0 · outbound
Large Language Models Hack Rewards, and Society Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8eea5c98-e98e-447e-8471-a3d9a23c1c4c · outbound
Large Language Models Hack Rewards, and Society Conference on Empirical Methods in Natural Language Processing , year=
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a023afe-25e1-47b6-989f-e0237c8e63be · outbound
Large Language Models Hack Rewards, and Society Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6546ecd0-79b7-457b-8cb9-68acc82e76e3 · outbound
Large Language Models Hack Rewards, and Society Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e62de06-0edc-4f60-866c-1d25240d7d1f · outbound
Large Language Models Hack Rewards, and Society Efficient memory management for large language model serving with
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc6291f-ad3a-456a-9a5d-0e34b35ed174 · outbound
Large Language Models Hack Rewards, and Society Artificial Intelligence and Law , year=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41dcb44f-03ae-4f4f-a95f-0e0edbacdee8 · outbound
Large Language Models Hack Rewards, and Society Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a45b5d8-65b1-432d-9b0d-25da7a571b89 · outbound
Large Language Models Hack Rewards, and Society Artificial Life , volume=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dbdf55d-109e-4ade-a36c-a1a34b1c19ee · outbound
Large Language Models Hack Rewards, and Society Categorizing variants of
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b4e23d-d159-438f-a465-a7cd95822a0c · outbound
Large Language Models Hack Rewards, and Society Fine-Tuning Language Models from Human Preferences
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e94129f-99c1-43dc-b5eb-4819b7e238ba · outbound
Large Language Models Hack Rewards, and Society Transactions on Machine Learning Research , year=
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8357985a-befe-4cf5-acfb-d58ffb1a555b · outbound
Large Language Models Hack Rewards, and Society International Conference on Machine Learning , year=
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 066e16c3-29bb-4dd0-8ef1-cdb84634eff6 · outbound
Large Language Models Hack Rewards, and Society Joint European conference on machine learning and knowledge discovery in databases , pages=
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb0636b4-d0cc-44f8-a2db-478e03d02a7e · outbound
Large Language Models Hack Rewards, and Society 2008 , publisher=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22b9cc5d-0635-44cf-be67-8fbffd195530 · outbound
Large Language Models Hack Rewards, and Society Journal of Banking & Finance , volume=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb146931-4c8a-43ee-b05e-707bc7ad60bf · outbound
Large Language Models Hack Rewards, and Society 2017 ieee symposium on security and privacy (sp) , pages=
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d3cb87f-dd97-4014-be20-5b22bb76bcb9 · outbound
Large Language Models Hack Rewards, and Society Proceedings of the national academy of sciences , volume=
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c92f32e6-971b-4d1f-aae4-fcea7f15585d · outbound
Large Language Models Hack Rewards, and Society The Quarterly Journal of Economics , volume=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75ae0e75-9a6e-4d43-b03d-2a1de4c82b93 · outbound
Large Language Models Hack Rewards, and Society 2011 , publisher=
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 854eab67-52f8-4ce1-91d1-788c5cf33e5f · outbound
Large Language Models Hack Rewards, and Society IEEE Transactions on Software Engineering , volume=
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8a62c8c-423f-431f-a4dd-870060063d99 · outbound
Large Language Models Hack Rewards, and Society Advances in neural information processing systems , volume=
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19c1af63-037c-46cf-9f48-b89d7de7acef · outbound
Large Language Models Hack Rewards, and Society International conference on foundations of software technology and theoretical computer science , pages=
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bcc912b-3254-4736-9c87-7b21ff450b35 · outbound
Large Language Models Hack Rewards, and Society ACM Computing Surveys , volume=
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a1ae977-4cf2-433a-9b2a-32ce414c877e · outbound
Large Language Models Hack Rewards, and Society ACM Computing Surveys (CSUR) , volume=
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c70e1219-0983-4038-8c46-e39919b6a405 · outbound
Large Language Models Hack Rewards, and Society Advances in neural information processing systems , volume=
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67f54172-0119-4fef-8d09-c76d2c08939a · outbound
Large Language Models Hack Rewards, and Society TradingAgents: Multi-Agents LLM Financial Trading Framework
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba3d1739-5bfe-48e8-a4e8-57442e83be39 · outbound
Large Language Models Hack Rewards, and Society Findings of the Association for Computational Linguistics: ACL 2024 , pages=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0b76931-b613-41f8-a053-5128de6b9ac5 · outbound
Large Language Models Hack Rewards, and Society Political Analysis , volume=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ad8bde6-bf97-4eda-b8fc-2c3f85f9c586 · outbound
Large Language Models Hack Rewards, and Society AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb9b8cd8-ff78-489a-b0f7-9da875eb5642 · outbound
Large Language Models Hack Rewards, and Society Findings of the Association for Computational Linguistics: EMNLP 2024 , pages=
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d17cf83-4421-420e-86b0-585a22b214fc · outbound
Large Language Models Hack Rewards, and Society Science Advances , volume=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 255e9924-10d8-406f-8d6c-6455072278fa · outbound
Large Language Models Hack Rewards, and Society Generative Language Models and Automated Influence Operations: Emerging Threats and Potential Mitigations
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e361583-db4b-4110-8795-6bf517c122d8 · outbound
Large Language Models Hack Rewards, and Society Navigating the Risks: A Survey of Security, Privacy, and Ethics Threats in LLM-Based Agents
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e33b3c3-f666-429f-b3a8-0f14e9ecc732 · outbound
Large Language Models Hack Rewards, and Society Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ee9eaa9-831d-41c4-8411-ffeff97c5c48 · outbound
Large Language Models Hack Rewards, and Society 2024 , eprint=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01738f1b-68fc-4a16-b4e5-74b2db40a2ce · outbound
Large Language Models Hack Rewards, and Society Public Administration Review , volume=
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b310ca2-d417-4290-a21d-3ddaeb7c9b1e · outbound
Large Language Models Hack Rewards, and Society American sociological review , volume=
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f87a040c-ddc9-495d-9ef1-53b93b73626d · outbound
Large Language Models Hack Rewards, and Society New York: Russell Sage Foundation , year=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02022335-1413-419c-9494-6bb26874e09a · outbound
Large Language Models Hack Rewards, and Society short-termism
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f82fd7d-0e14-4daf-b20b-0b5ac41f2e5d · outbound
Large Language Models Hack Rewards, and Society Monetary theory and practice: The UK experience , pages=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e52f5f1-3b63-4760-8fe1-e88529dc5060 · outbound
Large Language Models Hack Rewards, and Society RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73bf1c18-c609-4837-87ed-c7dfba1e7e44 · outbound
Large Language Models Hack Rewards, and Society DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb18a35c-3c6f-4fad-8258-7d320d3709c0 · outbound
Large Language Models Hack Rewards, and Society Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8624b820-36b4-4fcf-8606-1021e9b205ae · outbound
Large Language Models Hack Rewards, and Society A Long Way to Go: Investigating Length Correlations in RLHF
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc16fbca-74f8-41dc-90dd-e3b490d9c523 · outbound
Large Language Models Hack Rewards, and Society Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a821e81f-971b-4b59-8eea-821bd0920e4d · outbound
Large Language Models Hack Rewards, and Society Natural Emergent Misalignment from Reward Hacking in Production RL
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e531e21-b025-4d4f-b37c-bb78efe65315 · outbound
Large Language Models Hack Rewards, and Society Advances in Neural Information Processing Systems , volume=
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 275e842f-49d1-46a3-9022-42411fff8769 · outbound
Large Language Models Hack Rewards, and Society Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e47899e1-f552-4569-9203-bb1ba7c8a990 · outbound
Large Language Models Hack Rewards, and Society Learning to Discover at Test Time
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f84771c3-3d0f-43a8-be0f-38c68c4f85f2 · outbound
Large Language Models Hack Rewards, and Society Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcce7423-9742-4dff-ab3d-f34fc17c9a39 · outbound
Large Language Models Hack Rewards, and Society Algorithmic Collusion by Large Language Models
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09ea372e-6966-4b61-88c7-46ddd4e38b8c · outbound
Large Language Models Hack Rewards, and Society Can AI expose tax loopholes? Towards a new generation of legal policy assistants
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a92c3ba6-6776-40c3-bb1e-8971b9910b20 · outbound
Large Language Models Hack Rewards, and Society arXiv preprint arXiv:2603.20281 , year=
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3236340e-b708-4bf3-a947-2e2b2deb9d45 · outbound
Large Language Models Hack Rewards, and Society The Twelfth International Conference on Learning Representations , year=
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b36843a0-d119-4598-84e5-d47493a012c9 · outbound
Large Language Models Hack Rewards, and Society arXiv preprint arXiv:2507.08068 , year=
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e7d8c9d-acb6-4c4f-8d97-58ddd8ac2d83 · outbound
Large Language Models Hack Rewards, and Society Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b67b3ce-bc6b-4faf-977d-e3fb44c47ec3 · outbound
Large Language Models Hack Rewards, and Society Qwen3 Technical Report
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fca1f5f6-26be-4f95-bc46-a5bbed9aab11 · outbound
Large Language Models Hack Rewards, and Society 2025 , url =
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b01f78f-a34b-41a9-9a90-6ab7e0a4249e · outbound
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e2c5db9-676d-4e0e-b5b7-1995683c2eb2 · outbound
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d768157-7af4-4210-83a6-6e629150c6ed · outbound
Large Language Models Hack Rewards, and Society HealthBench: Evaluating Large Language Models Towards Improved Human Health
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef2fee43-d1f8-4bcb-9dcd-b053df353693 · outbound
Large Language Models Hack Rewards, and Society Richard and Koch, Gary G
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a0a8b9f-ffdd-4448-9cb7-15f7a2fce038 · outbound
Large Language Models Hack Rewards, and Society Emergent Misalignment : Narrow finetuning can produce broadly misaligned LLMs , May 2025
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a77f0fc-8f8c-492d-bf1f-8786a37c02b3 · outbound
Large Language Models Hack Rewards, and Society OpenAI GPT-5 System Card
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0eee1109-9cca-4592-a37b-6bb32e06bce9 · outbound
Large Language Models Hack Rewards, and Society Unresolved cited work
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1d78bf0-6b90-4310-9ad1-40dff0108329 · outbound
Large Language Models Hack Rewards, and Society The Fourteenth International Conference on Learning Representations , year=
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e9f0df3-28c3-40e4-8bf1-0ef793de5762 · outbound
Large Language Models Hack Rewards, and Society The International Conference on Learning Representations (ICLR) Blog Post Track , year=
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.