Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2406.14088.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:19:34.081749Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T18:48:49.536779Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 23c34f15-0efb-47fb-978c-a4d69d38317a · inbound
7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baf37d31-4399-40b3-9a93-4fdb43061efa · inbound
StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fda7b7fa-5ba2-4757-93b0-eeb1cbb426da · inbound
Reinforcement Learning Optimization for Large-Scale Learning: An Efficient and User-Friendly Scaling Library ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8ec1b39-c291-4b52-ba9d-2c16cba2ad62 · inbound
MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03e775ba-5b7d-4f06-8de8-c5101b999003 · inbound
Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f01dc95e-efa1-4994-8bca-1d0c052dc0f0 · inbound
TensorHub: Scalable and Elastic Weight Transfer for LLM RL Training ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 62b51473-9f4e-44f7-93fa-3a1685713e0c · inbound
AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cac5490b-9dab-4729-b291-ae7bfa2b0daa · inbound
PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ef2e30af-fb88-4ac4-929c-c88ea9e06416 · inbound
Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 06870b1d-8722-4fea-8bb3-1141e75b3305 · inbound
Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c920d508-64bc-4f6d-b9d8-261ba752fd9e · inbound
DynaResize: Runtime GPU Reallocation for Disaggregated LLM Post-Training ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.