Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.07927.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:18:54.577941Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T17:50:00.073152Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 992949d2-b9ae-474e-a18b-77742679f4f9 · inbound
Average-Reward Soft Actor-Critic Efficient Reinforcement Learning with Large Language Model Priors
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09abbb62-d7e7-4cbf-a5e4-31b3312d8702 · inbound
EVAL: EigenVector-based Average-reward Learning Efficient Reinforcement Learning with Large Language Model Priors
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 305b7bed-a869-4f01-8ebc-60d8f37803b9 · inbound
HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving Efficient Reinforcement Learning with Large Language Model Priors
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3983245-fbc8-40f6-96d3-144f483723f9 · inbound
Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration Efficient Reinforcement Learning with Large Language Model Priors
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 15f50169-fca5-4fc5-8757-d2b5660d27e8 · inbound
Reinforcement Learning Foundation Models Should Already Be A Thing Efficient Reinforcement Learning with Large Language Model Priors
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b8d68713-0cc0-4e3e-aea3-c6ddf33b5067 · inbound
LaGO: Latent Action Guidance for Online Reinforcement Learning Efficient Reinforcement Learning with Large Language Model Priors
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.