Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:02.986353Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2506.09202.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:02.986353Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-27T10:27:47.896922Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T17:20:00.517702Z
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c47bc253-0561-4887-8529-9a3809064937 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1c511b1-b97e-4a24-ac76-3583315dfd99 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 379f1434-9361-4d59-8b59-4ccbd6333370 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86f73a97-a64a-4d6b-9036-cd7c669c123f · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c94e70c-2d7b-46da-b108-9da3b240bee5 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Takeball There are four different rule-based policies, i-th policy will pick the i-th ball first
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2cadc6d3-6a9d-4f48-83b9-e7a07950a58f · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation adb002fe-3f00-4860-a89e-3e72b48120bc · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning These policies are corresponding to: No preference, prefer to go up first and prefer to go right first
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d36b757d-f311-4590-ad1b-b9c86dbe2772 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unsupervised Deep Embedding for Clustering Analysis
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d8c3380-d34b-49f4-9707-2f6ca7722d5e · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Behavior Regularized Offline Reinforcement Learning
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b92c03a-6ad1-47e1-9b08-ad2ba2ea496f · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Contrastive Clustering
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e6a45a0-9bee-4608-80df-30b917f176a3 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Deep Reinforcement Learning for Autonomous Driving: A Survey
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80ea89f9-9ee0-454e-bc69-aece7f06d8c5 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning Dataset Clustering for Improved Offline Policy Learning
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d86d253-c3c4-42db-ae9c-ef56f3c9b302 · outbound
Policy-Based Trajectory Clustering in Offline Reinforcement Learning URL http://dx.doi.org/10.1109/TNNLS.2023
Reference 2388
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 513d2eda-a545-4fcd-b8e7-42e07391a589 · inbound
Implicit Neural Representations of Individual Behavior Policy-Based Trajectory Clustering in Offline Reinforcement Learning
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7aa0886d-e7c1-4479-8a86-934233ff7c8d · inbound
ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning Policy-Based Trajectory Clustering in Offline Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.