Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-10T16:39:10.354308Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2607.07855.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-10T16:39:10.354308Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T04:15:52.407616Z
A source-named dated measurement, never combined with another source.
Source: cited_works
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a1f560ea-b84a-4e16-88d3-bd6708dbc0f3 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Option-aware temporally abstracted value for offline goal-conditioned reinforcement learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d4be5774-c99d-4413-b3b9-0b6f58dbc2df · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Let offline RL flow: Training conservative agents in the latent space of normalizing flows
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fce20aff-8b51-4315-84a9-4f6b4db921d5 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Hindsight experience replay
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9dc671b4-8ff8-48bc-a8b2-c9e199810a01 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Graph-assisted stitching for offline hierarchical reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2c1c59ce-c563-4f9c-97ae-1cd92da5cf6e · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Test-time offline reinforcement learning on goal-related experience
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1790222e-ff88-47b5-9318-fc931e3d15fe · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Flowpg: action-constrained policy gradient with normalizing flows.Advances in Neural Information Processing Systems, 36:20118–20132
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 63602eeb-a32e-435b-9602-9c720e090ebe · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Maximum entropy reinforcement learning via energy-based normalizing flow.Advances in Neural Information Processing Systems, 37:56136–56165, 2024
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 76551841-f9cf-482b-9d5c-bb2ea89eabec · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL NICE: Non-linear Independent Components Estimation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation eda767ff-090b-44a0-a44d-1922026a0618 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Density estimation using real NVP
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1fcabc6e-5df7-4320-a04b-cbb1c354431e · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Inference via interpolation: Contrastive representations provably enable planning and inference.Advances in Neural Information Processing Systems, 37:58901–58928, 2025
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3d58bd92-2834-4b98-8f86-b6f91dcedbcd · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Contrastive learning as goal-conditioned reinforcement learning.Advances in Neural Information Processing Systems, 35:35603– 35620, 2022
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 39c36489-9210-4231-a26f-b2c8ca5746a8 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Normalizing Flows are Capable Models for Continuous Control
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 242702d1-010e-4643-8d22-4a1e8561a49c · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Physics-informed value learner for offline goal-conditioned reinforcement learning.arXiv preprint arXiv:2509.06782, 2025
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 22a34649-58e6-475c-8033-d93cb4df91da · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Hierarchical entity- centric reinforcement learning with factored subgoal diffusion
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation defb4118-cf24-4249-aa4a-3916073cda9a · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Diffused task-agnostic milestone planner.Advances in Neural Information Processing Systems, 36:387–405, 2023
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 67740f3d-7a17-4a7e-8736-e0b1696b0b52 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Learning to reach goals via diffusion
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bef3d8e2-8000-435b-8d13-6bb3bacd1d9d · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Conservative offline goal-conditioned implicit V-learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a5214038-2b40-477f-81fc-f6aea0689b8e · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Glow: Generative flow with invertible 1x1 convolutions.Advances in neural information processing systems, 31
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7c85920c-f05f-4c4f-8f51-ef43474fb1b4 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline reinforcement learning with implicit q-learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e7554da2-c2c7-448f-b792-d7c9f3641539 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL State-covering trajectory stitching for diffusion planners
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation af6f7d4d-8aad-4988-bc45-eb5b6f2f2c38 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL GCHR : Goal-Conditioned Hindsight Regularization for Sample-Efficient Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8af1fe5e-7ce1-4568-9475-3ff63c78dd8a · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 392a92a7-0807-4b3f-927a-077d51a41167 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Metric residual network for sample efficient goal- conditioned reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a5c757b8-2e77-4d2c-a0bf-64270616b7da · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Generative Trajectory Stitching through Diffusion Composition
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9ce5ffd3-bfce-47df-b279-2e80dce16d4a · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL How Far I'll Go: Offline Goal-Conditioned Reinforcement Learning via $f$-Advantage Regression
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0dc06858-e9ae-4a47-b03c-9c0fff2cb4a8 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Horizon generalization in reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 44393bcf-7b5b-4dec-8498-ecea8f77737d · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline goal-conditioned reinforcement learning with quasimetric representations
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 40569da3-c1de-4928-b1c6-ae2b1a701111 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline goal-conditioned reinforcement learning with quasimetric representations
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f28d030e-dcb3-42f3-9520-c547166a530d · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2b058db8-25b2-4753-a69f-f0f2b3afcc3c · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Test-Time Graph Search for Goal-Conditioned Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 22e47378-f8c3-4f0f-b601-cc2f1a39f569 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Normalizing flows for probabilistic modeling and inference.Journal of Machine Learning Research, 22(57):1–64
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 65c985d7-65c7-4e3c-9504-09670d9ae058 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Ogbench: Benchmarking offline goal-conditioned rl
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d03087d1-d02b-48f3-91f3-a78da3beef59 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL HIQL: Offline goal-conditioned RL with latent states as actions
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c5a727c8-13b9-454a-9d4b-68673740704c · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Foundation Policies with Hilbert Representations
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c63ff626-42f6-4438-943e-66e6f8706a76 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Flow Q-Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f679789f-7874-4f84-b3c5-7ce43fff3acd · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e1b51905-8eec-40b9-a270-4101c89ebcc1 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Bridging offline reinforcement learning and imitation learning: A tale of pessimism
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f73311ca-3712-4836-9759-fb1fe89ffc1e · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Goal-Conditioned Imitation Learning using Score-based Diffusion Policies
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6431ec90-ccd5-48ed-af97-cc237cf713f6 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Score models for offline goal-conditioned reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e65c9b04-a47b-493b-aa93-bef4ebc0e224 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Parrot: Data-Driven Behavioral Priors for Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8cc1b3c7-3f1f-4a4a-816c-d699a0d8b37f · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL GOPlan: Goal- conditioned offline reinforcement learning by planning with learned models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7c16e409-2c3d-44a4-8b2a-c71258c8dad2 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Optimal goal-reaching reinforcement learning via quasimetric learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e3419ac9-02e0-41dc-ba5f-9a937b33a7d7 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Improving Exploration in Soft-Actor-Critic with Normalizing Flows Policies
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 09665de8-f94d-46a0-a3f4-80d692a5fa30 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL A policy-guided imitation approach for offline reinforcement learning.Advances in neural information processing systems, 35:4085–4098, 2022
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d3ad5a0c-21ba-4904-8f98-7ccc30e175dc · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL An optimal discriminator weighted imitation perspective for reinforcement learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f853da24-14b5-4be2-aeda-6893d6b98a26 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Breadth-first exploration on adaptive grid for reinforcement learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b8062c9b-78ec-48cd-9a59-525f9786f700 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Scaling goal-conditioned reinforcement learning with multistep quasimetric distances
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6952dee9-8e9a-49c8-91f4-b55444fe9579 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Flattening hierarchies with policy bootstrapping
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 250229ea-a3d1-4529-aef6-be33ef4dfb3b · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a552ccad-5955-4c31-924e-3f136da9190b · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a4dadcbf-c690-46a1-a6a0-bdad8c4494a1 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 49f18931-36a5-4330-b001-51413b1829ba · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f62d64d1-a2f9-4ca2-b1b3-ddf9754cee4b · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL ∆∗ = 0 iff w is on an optimal path
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 882866f3-3199-4e05-a5b2-a4c7a51efd0d · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL ELBO/ODE estimates.AWR requires evaluating logπ H(z|s, g) for the weighted MLE objective
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3c5906cd-6e4a-4294-b927-78de85192af1 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8132e8c2-e28c-4777-85a1-69e8c97f5748 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL valid corridor
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 27527219-e277-442d-8551-91a8c3ef2063 · outbound
NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL lucky exploration
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8a9a3960-da33-4586-a473-ad6e4c61c817 · inbound
DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.