Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2503.17682.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:34:55.643156Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 51847302-7e16-4658-a595-1427070a47c5 · inbound
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 418727c9-566c-4acd-87f1-fd81820d4ac0 · inbound
Generative RLHF-V: Learning Principles from Multi-modal Human Preference Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9942e338-8dd3-41aa-b0c4-7e0716be7eab · inbound
USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b05dc424-5d68-43ba-ac7a-a20c44e3dc4b · inbound
The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0008580-4fda-476d-8da8-e57453f210b6 · inbound
HKGAI-V1: Towards Regional Sovereign Large Language Model for Hong Kong Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20002b8-7922-43f0-b393-b2d81a1ec7c8 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 296
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4857d8f9-2705-4462-bcc1-a59c1bc3add8 · inbound
Policy Gradient Primal-Dual Method for Safe Reinforcement Learning from Human Feedback Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8c3355a2-36aa-4c0f-b4a6-16f7423df480 · inbound
You Snooze, You Lose: Automatic Safety Alignment Restoration through Neural Weight Translation Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cdf1b522-e609-4f2c-9da6-e75c104e5caf · inbound
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12463d47-14de-4efb-95ad-283da661d8d5 · inbound
Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1953ddf-e0f4-47b4-bff3-cbaac8f1012e · inbound
Decoding Multimodal Cues: Unveiling the Implicit Meaning Behind Hateful Videos Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b3aab78-bba6-4d2b-a88a-1c883c5ad23c · inbound
Enhancing LLMs through human feedback: a journey towards self-improvement Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eb58766-c462-461b-9165-92440e6ca3cf · inbound
RoCo-ACE: Rollout-Conditioned Online Distillation for Retention-Aware Knowledge Injection Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.