Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:09:19.727723Z
Paper Citation Record · LEDGER
As of 1 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2606.25442.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-25T21:09:19.727723Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ef2b4891-04f4-4c96-a460-7e69a64d9593 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 38dfe6e5-ba1c-4977-afac-63d06655746c · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 00540679-92bc-4eb7-90b0-b319da0f40a2 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models A comprehensive survey on the trustworthiness of large language models in healthcare.arXiv preprint arXiv:2502.15871, 4
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ee84a894-fcc4-4c2f-8167-c6393d390948 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 2b841dd2-ce3d-42d2-bb37-1935f1511398 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 16bbce05-05dd-469f-a85c-eaa325940e29 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Finetuned Language Models Are Zero-Shot Learners
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 30042e0e-8060-4138-b776-f01763a18085 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Advances in neural information processing systems , volume=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 063af3bb-64f1-4ef0-9510-da9ca0dc6c30 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 33869db0-c45f-4fc3-8ac6-24afb875e31e · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Advances in neural information processing systems , volume=
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eb0da7c-0746-4fe1-bd1b-b1ec72a2ccec · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 9ed7fa22-9d02-43eb-8973-42746823784a · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Advances in neural information processing systems , volume=
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12d7ace-971e-4ab8-b91e-aeaccb56cfc6 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2023 , number =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7802f74d-0b69-40cc-901c-f6c0a5117119 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2021 , url =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 424f573e-5625-465f-a56e-6103c5d7eb40 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac2c9878-a75b-4fe0-b6e8-1624c4d210da · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2021 , url =
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31646f2f-037b-4c44-82c5-fe895853b948 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Constitutional AI: Harmlessness from AI Feedback
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T07:38:15.114855+00:00.
Observation 8e26d784-e1d1-4784-a47d-feedb03f575d · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Deliberative Alignment: Reasoning Enables Safer Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0c2cfc4a-c5e1-42aa-8778-d54b4fe00fc1 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models On-Policy Context Distillation for Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5e1ced3a-7f5d-4991-a6c0-bcdb79932f1e · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation e10b7495-3954-4e47-a9bb-616d289a5e6e · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Safety Tax: Safety Alignment Makes Your Large Reasoning Models Less Reasonable
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 8ee96c02-868e-42c1-8bdd-ef45521edad6 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Proceedings of ICLR , year=
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39c5b70-4939-4ee7-86e7-2ddfe1e12ba7 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models The twelfth international conference on learning representations , year=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66bc25e0-c54f-4271-b3f3-88da543a8dfa · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2025 , month =
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0cd011e-8d8b-4e24-a5cf-6bea362d161d · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2024 , eprint=
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab8c767b-04bb-4f45-a8c8-973efd854090 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2024 , eprint=
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd28c2e-5b09-47c3-9cee-e94441678bbf · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2025 , eprint=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbc3ed8f-f90e-4dd9-9fdc-e053bf2fa2ef · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation b2b9d4d1-78bb-43ad-a298-0b44dac91934 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2023 , eprint=
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52df4b8f-95b3-4843-b44f-151ceebbe86a · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Let's Verify Step by Step
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation f31f60cf-3d9a-457e-be1f-7e86f176ce46 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models XSTest: A test suite for identifying exaggerated safety behaviours in large language models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 6a3b0c63-c3a4-4c76-b368-32fbc6508cd2 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e58cc4d2-add5-44a0-8e57-a1917d85d278 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Advances in Neural Information Processing Systems , volume=
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae156d65-22ee-4fa7-a915-5116aeb5f20f · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Don’t let the claw grip your hand: A security analysis and defense framework for OpenClaw
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation c9a31ea6-4098-4c20-b102-8913135f6e15 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models IEEE Transactions on Pattern Analysis and Machine Intelligence , year=
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39b53d08-9530-4e02-a025-20c1613b9827 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Proceedings of the 26th annual international conference on machine learning , pages=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebde2d01-7dd4-4c44-b1b9-86b91501ab39 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac0dbfff-3d16-4027-9c79-e09c893b51e6 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6b0741c-e2bd-49ea-86e1-b48c43f2d4d4 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Six Musts and Six Don'ts
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1199e169-126d-4f67-84be-b8161cb526aa · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ef722e4-291e-4598-b0ab-a91b14451974 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models ArXiv , year=
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1575c0b-db97-42d4-a424-7ccff2530cb6 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models AlphaAlign: Incentivizing Safety Alignment with Extremely Simplified Reinforcement Learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5300cfbe-bb71-417c-a802-6545904e7933 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Mitigating the safety alignment tax with null-space constrained policy optimization.arXiv preprint arXiv:2512.11391
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation a240646d-925d-4c5d-8573-fc7eca196ad5 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 9b5ea55a-911f-4b47-aa53-578cc0a2dbb6 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Qwen3Guard Technical Report
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0d4eab1e-2d4b-43b0-981f-b8180462b169 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models GPT-4o System Card
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5424a20c-1c24-40d1-9f2d-1354db667f49 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Advances in neural information processing systems , volume=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36b75ef0-39c3-4866-96ea-7963ec50229a · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models CARES: Comprehensive Evaluation of Safety and Adversarial Robustness in Medical LLMs
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 520ee891-6745-4005-90f6-9693777cd455 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Applied Sciences , volume=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ac1b6bf-d013-413f-8a74-311da2741bd9 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Conference on health, inference, and learning , pages=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9371e08-57b8-40ce-b1ce-9aaf00f74c1e · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track , year=
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 336bcbc7-9e21-49cf-8bc1-50df8dd86755 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 376865d9-67da-4068-b0ca-abcb4744e035 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Lexam: Benchmarking legal reasoning on 340 law exams
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 67f19427-f5e0-441e-b2e8-8d8d7016ce46 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models FinanceBench: A New Benchmark for Financial Question Answering
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation b635e296-c506-4f49-82fb-bdda3f0854f7 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbbc50e2-44d1-4fc9-a53e-dc720838c09c · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2025 , url=
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6e2df9-644e-499a-a87c-90ac22b74fb1 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2025 , eprint=
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51f8ef4-1671-4044-bc7a-8aa6f0157764 · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2026 , journal =
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ac1baa8-2446-4e9d-a865-9dd4ab8c0a6d · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2026 , eprint=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cedabc4c-819e-46f1-84bb-7b2fe63782fd · outbound
PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models 2024 , month = jul, doi =
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.