Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:18:42.180146Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2506.22008.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:18:42.180146Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3d7dec4f-6150-4ed1-994b-60799a054b25 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning It is an open-world city simulation originally proposed by Sestini et al
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8fd18dfc-2dfc-47e6-bd2e-4db9c7b22d91 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Semi-supervised reward learning for offline reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 712750b4-bcd7-4a39-903d-6e8105fc559e · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Imitating Human Behaviour with Diffusion Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db243e78-9b73-4234-934d-fe74ed65f3cc · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b91c065-8c05-4f75-97eb-340212b64dd2 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Real-Time Diffusion Policies for Games: Enhancing Consistency Policies with Q-Ensembles
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bfbfab27-e9ab-46be-8d03-b18b17121187 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Offline Learning from Demonstrations and Unlabeled Experience
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0fc60be-e3a7-473b-a2d5-a12719025707 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Offline Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf7933e9-9537-4676-b198-4e8b16331d02 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning As we describe in Section 3, it is especially suitable for offline settings
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 42e8ad3a-8bb7-4dbd-a2bf-00e9e752eedc · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning All of these approaches require optimal expert demonstrations – that is, demonstrations generated by an optimal policy
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 23f83ca3-7541-496e-b363-2e8a34dc3744 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning An episode is marked a success if the agent reaches the goal before the timeout
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 59acd6bf-03da-4eaa-96a4-29a66e2f455c · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning We provide more details about the baselines in Section 4.2 of the main paper
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f204ddd1-e672-44b6-95c5-8bafe6d88323 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d4c60b3a-dd2e-4220-9dd3-04619ddab4d2 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Efficient Active Imitation Learning with Random Network Distillation
Reference 1995
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0468167d-bdc4-4e3a-8b6a-eb989e5aefb8 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning The Provable Benefits of Unsupervised Data Sharing for Offline Reinforcement Learning
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31bad089-20d2-4f3c-b523-8c9ce563d6b3 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Technical challenges of deploying reinforcement learning agents for game testing in aaa games
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0f97daf-17e3-4e84-9b5d-5f5a97969c03 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Demonstration-efficient inverse reinforcement learning in procedurally generated environments
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 213c6f6d-44c1-4820-bc60-a440fd8fe231 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2209042a-0467-4f47-87c4-06c0de714503 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Towards informed design and validation assistance in computer games using imitation learn- ing
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3ebf8c51-d315-40b4-bddf-d33506d93911 · outbound
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning Beyond Reward: Offline Preference-guided Policy Optimization
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.