Pith. sign in

Paper Citation Record · LEDGER

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL

As of 17 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 1 inbound Pith citation observation for arXiv:2607.07855.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07855 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T16:39:10.354308Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T04:15:52.407616Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact16
  • verified fuzzy33
  • unresolved5
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1f560ea-b84a-4e16-88d3-bd6708dbc0f3 · outbound

This paper cites Option-aware temporally abstracted value for offline goal-conditioned reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Option-aware temporally abstracted value for offline goal-conditioned reinforcement learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-10T16:47:24.447006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:4174fe798ef4daf5f73332653574ba44f9658cb8a6233a2e104cadcd7f72f860

Observation d4be5774-c99d-4413-b3b9-0b6f58dbc2df · outbound

This paper cites Let offline RL flow: Training conservative agents in the latent space of normalizing flows.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Let offline RL flow: Training conservative agents in the latent space of normalizing flows

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.786980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:7779edf33a6d33701d428d68858371559416429a6e407587a4adf8029d0f85f9

Observation fce20aff-8b51-4315-84a9-4f6b4db921d5 · outbound

This paper cites Hindsight experience replay.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Hindsight experience replay

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.794348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:981520daedbb69229697e69f74ff298288c64ecf54b2e93b5ec55776cb1e292a

Observation 9dc671b4-8ff8-48bc-a8b2-c9e199810a01 · outbound

This paper cites Graph-assisted stitching for offline hierarchical reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Graph-assisted stitching for offline hierarchical reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.782508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:aac4ca940d01a41891cb277937224e2986098898b2e48b8ac2c675bd0d1f1cbb

Observation 2c1c59ce-c563-4f9c-97ae-1cd92da5cf6e · outbound

This paper cites Test-time offline reinforcement learning on goal-related experience.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Test-time offline reinforcement learning on goal-related experience

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.784706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:2bb3bbe28ff9172f99916459fa4f795e149f9ee7c882cf4635f4aa43b9c010c4

Observation 1790222e-ff88-47b5-9318-fc931e3d15fe · outbound

This paper cites Flowpg: action-constrained policy gradient with normalizing flows.Advances in Neural Information Processing Systems, 36:20118–20132.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Flowpg: action-constrained policy gradient with normalizing flows.Advances in Neural Information Processing Systems, 36:20118–20132

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.814651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:f1c3f79fc57819ae6921e15195508ff101573c936a894fc3b92186ae7936ad29

Observation 63602eeb-a32e-435b-9602-9c720e090ebe · outbound

This paper cites Maximum entropy reinforcement learning via energy-based normalizing flow.Advances in Neural Information Processing Systems, 37:56136–56165, 2024.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Maximum entropy reinforcement learning via energy-based normalizing flow.Advances in Neural Information Processing Systems, 37:56136–56165, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.796704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:613983f3ab1c716885cb25d9abdfc2d9a691af4afbf777c012856596dfc52ad8

Observation 76551841-f9cf-482b-9d5c-bb2ea89eabec · outbound

This paper cites NICE: Non-linear Independent Components Estimation.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL NICE: Non-linear Independent Components Estimation

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.446897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:e01ecbceec43c57abe928fb0ed02d521c4ea3ffd3a5126e311f27c906e14b01e

Observation eda767ff-090b-44a0-a44d-1922026a0618 · outbound

This paper cites Density estimation using real NVP.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Density estimation using real NVP

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.832686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:f4762fbe2ba069c00f367ab3d04a52e718ec3f00adcfdb04f4a94397130c7f90

Observation 1fcabc6e-5df7-4320-a04b-cbb1c354431e · outbound

This paper cites Inference via interpolation: Contrastive representations provably enable planning and inference.Advances in Neural Information Processing Systems, 37:58901–58928, 2025.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Inference via interpolation: Contrastive representations provably enable planning and inference.Advances in Neural Information Processing Systems, 37:58901–58928, 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.855972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:6f5894049505ed1d6901d6dad5ee76f291d016eecf355de8730642cd157628fc

Observation 3d58bd92-2834-4b98-8f86-b6f91dcedbcd · outbound

This paper cites Contrastive learning as goal-conditioned reinforcement learning.Advances in Neural Information Processing Systems, 35:35603– 35620, 2022.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Contrastive learning as goal-conditioned reinforcement learning.Advances in Neural Information Processing Systems, 35:35603– 35620, 2022

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.853939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:604589d69d08e907f3d7e0ed65133da97fc11249d51a6426cfa3765999c2fb92

Observation 39c36489-9210-4231-a26f-b2c8ca5746a8 · outbound

This paper cites Normalizing Flows are Capable Models for Continuous Control.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Normalizing Flows are Capable Models for Continuous Control

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.462029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:58e7b80ac33357d3227c6b4a92e2e6b7715e64a2eb75e8494460941788034115

Observation 242702d1-010e-4643-8d22-4a1e8561a49c · outbound

This paper cites Physics-informed value learner for offline goal-conditioned reinforcement learning.arXiv preprint arXiv:2509.06782, 2025.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Physics-informed value learner for offline goal-conditioned reinforcement learning.arXiv preprint arXiv:2509.06782, 2025

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-10T16:47:24.464996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:e15ccde509fc46a3d59c47ac36d8e230eb0e3d37b816a434917390b465a4452f

Observation 22a34649-58e6-475c-8033-d93cb4df91da · outbound

This paper cites Hierarchical entity- centric reinforcement learning with factored subgoal diffusion.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Hierarchical entity- centric reinforcement learning with factored subgoal diffusion

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.854139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:2293982380a346d4af5a19dbfd5d144c9b748a524938f5cd94014b965d8a1d45

Observation defb4118-cf24-4249-aa4a-3916073cda9a · outbound

This paper cites Diffused task-agnostic milestone planner.Advances in Neural Information Processing Systems, 36:387–405, 2023.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Diffused task-agnostic milestone planner.Advances in Neural Information Processing Systems, 36:387–405, 2023

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.859775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:07e54d798a83c5e6778ff0da22ee9a1a449b3aefc63bbc0bc43b7bd5533ccf96

Observation 67740f3d-7a17-4a7e-8736-e0b1696b0b52 · outbound

This paper cites Learning to reach goals via diffusion.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Learning to reach goals via diffusion

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.843607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:a00292f2537a2c4c7f14a20b3eb35c92fa7aeceeb337a255e83d08869c25aa41

Observation bef3d8e2-8000-435b-8d13-6bb3bacd1d9d · outbound

This paper cites Conservative offline goal-conditioned implicit V-learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Conservative offline goal-conditioned implicit V-learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.844113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:5057500c2e2bc3b139808d9c69f0cf925aaf03df940e644b373dbb018c542f55

Observation a5214038-2b40-477f-81fc-f6aea0689b8e · outbound

This paper cites Glow: Generative flow with invertible 1x1 convolutions.Advances in neural information processing systems, 31.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Glow: Generative flow with invertible 1x1 convolutions.Advances in neural information processing systems, 31

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.840355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:bd5724029f435f3cb71156ad41843383ab03e1b1eed915aa473ac05459b07ad9

Observation 7c85920c-f05f-4c4f-8f51-ef43474fb1b4 · outbound

This paper cites Offline reinforcement learning with implicit q-learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline reinforcement learning with implicit q-learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.838043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:a525620f5e46d306363a4f19b45c12a7c80816cf8c306dc48ab8275f216544b8

Observation e7554da2-c2c7-448f-b792-d7c9f3641539 · outbound

This paper cites State-covering trajectory stitching for diffusion planners.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL State-covering trajectory stitching for diffusion planners

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.780750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:cd78225325a265e31504c8a2a4dc6a1842f0fa7d92f4335190fbbb639638b0b7

Observation af6f7d4d-8aad-4988-bc45-eb5b6f2f2c38 · outbound

This paper cites GCHR : Goal-Conditioned Hindsight Regularization for Sample-Efficient Reinforcement Learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL GCHR : Goal-Conditioned Hindsight Regularization for Sample-Efficient Reinforcement Learning

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.434557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:c4687252d7c32d07bcbf846be7eb41640766f4e881cae2d415ca744c8821475c

Observation 8af1fe5e-7ce1-4568-9475-3ff63c78dd8a · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.456442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:1087910bfc36fd93ceaa6edd5fe67548281c42a3b017092a78212784b1c37300

Observation 392a92a7-0807-4b3f-927a-077d51a41167 · outbound

This paper cites Metric residual network for sample efficient goal- conditioned reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Metric residual network for sample efficient goal- conditioned reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.845961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:15d359bf61c050643eb64374c48572b27587e3475346f35ac43394bed1681545

Observation a5c757b8-2e77-4d2c-a0bf-64270616b7da · outbound

This paper cites Generative Trajectory Stitching through Diffusion Composition.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Generative Trajectory Stitching through Diffusion Composition

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T16:47:24.431975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:8b2bef81a337264c83fafbc40b080c915dd75bb9f81c772c3c47e60703ce1940

Observation 9ce5ffd3-bfce-47df-b279-2e80dce16d4a · outbound

This paper cites How Far I'll Go: Offline Goal-Conditioned Reinforcement Learning via $f$-Advantage Regression.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL How Far I'll Go: Offline Goal-Conditioned Reinforcement Learning via $f$-Advantage Regression

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.470362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:d0efdf1f72b9e1c305e51df392ed34a19dcdc70248dd16abcb5f5dadace6db66

Observation 0dc06858-e9ae-4a47-b03c-9c0fff2cb4a8 · outbound

This paper cites Horizon generalization in reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Horizon generalization in reinforcement learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.847763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:4defb014506e5a94894aaa982ed2b99188a7219a0cc72df073d31056b72b5809

Observation 44393bcf-7b5b-4dec-8498-ecea8f77737d · outbound

This paper cites Offline goal-conditioned reinforcement learning with quasimetric representations.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline goal-conditioned reinforcement learning with quasimetric representations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.852325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:3c022d4af43a0e71d10d2333553c48d7c1c4f6ed5d6574620aa87b276826cb8f

Observation 40569da3-c1de-4928-b1c6-ae2b1a701111 · outbound

This paper cites Offline goal-conditioned reinforcement learning with quasimetric representations.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Offline goal-conditioned reinforcement learning with quasimetric representations

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.839921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:d1a7375f73078423fff4785c2d11a01d1aca3a71ae696b36dbd96172910751f7

Observation f28d030e-dcb3-42f3-9520-c547166a530d · outbound

This paper cites Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.463502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:94a491e4a754cedc8b9cb43d388670a579095456d6879b15f2026c305b566b27

Observation 2b058db8-25b2-4753-a69f-f0f2b3afcc3c · outbound

This paper cites Test-Time Graph Search for Goal-Conditioned Reinforcement Learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Test-Time Graph Search for Goal-Conditioned Reinforcement Learning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.465922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:8c7b95a21326eceea02d18cb74b6939c9363d872a6951eae3261bcf73da38729

Observation 22e47378-f8c3-4f0f-b601-cc2f1a39f569 · outbound

This paper cites Normalizing flows for probabilistic modeling and inference.Journal of Machine Learning Research, 22(57):1–64.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Normalizing flows for probabilistic modeling and inference.Journal of Machine Learning Research, 22(57):1–64

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.858017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:618388db36e34fb945847afe8c862a572a7b658fb7cec37db38e74b827a0685c

Observation 65c985d7-65c7-4e3c-9504-09670d9ae058 · outbound

This paper cites Ogbench: Benchmarking offline goal-conditioned rl.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Ogbench: Benchmarking offline goal-conditioned rl

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.861789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:7f63b0fd55f7dcc3ae446007d73cf603a9c4a6490ed09f64acb2f7fe6d41bccc

Observation d03087d1-d02b-48f3-91f3-a78da3beef59 · outbound

This paper cites HIQL: Offline goal-conditioned RL with latent states as actions.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL HIQL: Offline goal-conditioned RL with latent states as actions

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.841856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:280bdfa4ee11646a3ed5877195f7b358b3a3e1def76da9f5a6164ca9c50ac809

Observation c5a727c8-13b9-454a-9d4b-68673740704c · outbound

This paper cites Foundation Policies with Hilbert Representations.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Foundation Policies with Hilbert Representations

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.472888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:37ede01274860f98b6e4c2443aa626b9ac142eeb3e4cf8036bdc1ae866d85ac0

Observation c63ff626-42f6-4438-943e-66e6f8706a76 · outbound

This paper cites Flow Q-Learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Flow Q-Learning

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.449484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:7149d1516e8de47e4651e0550658a11cd3a4f7ada08cc0c0a78542187868b982

Observation f679789f-7874-4f84-b3c5-7ce43fff3acd · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.467632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:b9a52dd1317ee14e93cb64f4a1e730792233086b8fc942c440fec5bd3d862e72

Observation e1b51905-8eec-40b9-a270-4101c89ebcc1 · outbound

This paper cites Bridging offline reinforcement learning and imitation learning: A tale of pessimism.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Bridging offline reinforcement learning and imitation learning: A tale of pessimism

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.855688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:11277e4fe96e57fdf055f704b0378cbd87fc4d70bc39af95cbf6172d13f396aa

Observation f73311ca-3712-4836-9759-fb1fe89ffc1e · outbound

This paper cites Goal-Conditioned Imitation Learning using Score-based Diffusion Policies.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Goal-Conditioned Imitation Learning using Score-based Diffusion Policies

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.459306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:f5ad6d552b3a78d3af71cb7e3704c4a7568dbf23ba8a7b7e2545e53c17cf751a

Observation 6431ec90-ccd5-48ed-af97-cc237cf713f6 · outbound

This paper cites Score models for offline goal-conditioned reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Score models for offline goal-conditioned reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.827069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:0fc2c6129d45cd33bc9e4bfb08185f99c5f0e3e58569f7860ab3618f3b1a4f99

Observation e65c9b04-a47b-493b-aa93-bef4ebc0e224 · outbound

This paper cites Parrot: Data-Driven Behavioral Priors for Reinforcement Learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Parrot: Data-Driven Behavioral Priors for Reinforcement Learning

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.454982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:4019834f8211227581de1b7921ba1d00f405f8acf9e9b778009d621e3834e0b8

Observation 8cc1b3c7-3f1f-4a4a-816c-d699a0d8b37f · outbound

This paper cites GOPlan: Goal- conditioned offline reinforcement learning by planning with learned models.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL GOPlan: Goal- conditioned offline reinforcement learning by planning with learned models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.827971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:3487b8ef222a8b2116cc5f3a78fac5e0ea552805881351e2dcba44ff82142914

Observation 7c16e409-2c3d-44a4-8b2a-c71258c8dad2 · outbound

This paper cites Optimal goal-reaching reinforcement learning via quasimetric learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Optimal goal-reaching reinforcement learning via quasimetric learning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.815618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:c0eb80ca5565b990cc9da3c24132b782f4f11d0439fab95c894dede490012b13

Observation e3419ac9-02e0-41dc-ba5f-9a937b33a7d7 · outbound

This paper cites Improving Exploration in Soft-Actor-Critic with Normalizing Flows Policies.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Improving Exploration in Soft-Actor-Critic with Normalizing Flows Policies

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:47:24.470801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:7e3d20ee4d780cc0fd9b96e769397982b9bf352ef7e357138fbd6219b3044d23

Observation 09665de8-f94d-46a0-a3f4-80d692a5fa30 · outbound

This paper cites A policy-guided imitation approach for offline reinforcement learning.Advances in neural information processing systems, 35:4085–4098, 2022.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL A policy-guided imitation approach for offline reinforcement learning.Advances in neural information processing systems, 35:4085–4098, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.810526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:240d75498672ec8a7d75be7fc5a128406b71711776fda3051c51993925222f47

Observation d3ad5a0c-21ba-4904-8f98-7ccc30e175dc · outbound

This paper cites An optimal discriminator weighted imitation perspective for reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL An optimal discriminator weighted imitation perspective for reinforcement learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.824630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:8d3c5de3d7de2f5fbebc489023aed947709426dcf19d0a75ab44853ef5e13847

Observation f853da24-14b5-4be2-aeda-6893d6b98a26 · outbound

This paper cites Breadth-first exploration on adaptive grid for reinforcement learning.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Breadth-first exploration on adaptive grid for reinforcement learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.804150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:99fa12c9360c358be98f96d182538d281740d04cb607d999559d1fbcecbc2f62

Observation b8062c9b-78ec-48cd-9a59-525f9786f700 · outbound

This paper cites Scaling goal-conditioned reinforcement learning with multistep quasimetric distances.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Scaling goal-conditioned reinforcement learning with multistep quasimetric distances

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.829050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:17613b505b85324a9b53e0414281425c46ddafbca65b70bfc87b19163a37236b

Observation 6952dee9-8e9a-49c8-91f4-b55444fe9579 · outbound

This paper cites Flattening hierarchies with policy bootstrapping.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Flattening hierarchies with policy bootstrapping

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-10T16:47:24.468473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:f1f1c3ee731bd07d2c729063eab6213207ed8ce1ac86ed36f57a44e3afb0c4ab

Observation 250229ea-a3d1-4529-aef6-be33ef4dfb3b · outbound

This paper cites an unresolved cited work.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-07-10T16:47:24.798899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:046046f663cd00cdeb82bca4062fb9b4106f07a0ad2e0db154c77ff5b01aa51b

Observation a552ccad-5955-4c31-924e-3f136da9190b · outbound

This paper cites an unresolved cited work.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-07-10T16:47:24.834541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:df3d55edc4669f573d6ac5bda45ebc3723b12e39f9ac6873f06f91d5d40470a7

Observation a4dadcbf-c690-46a1-a6a0-bdad8c4494a1 · outbound

This paper cites an unresolved cited work.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-07-10T16:47:24.830859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:9df030fab408c3d539e3f70842f21a63d921d0f26fe4f2c2bb6b5cd1a69afd84

Observation 49f18931-36a5-4330-b001-51413b1829ba · outbound

This paper cites an unresolved cited work.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-07-10T16:47:24.800887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:568466e21fc8b9e36df3897f6111948a667ffe4ef36036a274727fecec5b350b

Observation f62d64d1-a2f9-4ca2-b1b3-ddf9754cee4b · outbound

This paper cites ∆∗ = 0 iff w is on an optimal path.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL ∆∗ = 0 iff w is on an optimal path

Reference 53

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T16:47:24.807957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:20751cc270d88d2115c67dfc265ee0f6fbcacb5c9e1a56081e5e4bd45b4dd644

Observation 882866f3-3199-4e05-a5b2-a4c7a51efd0d · outbound

This paper cites ELBO/ODE estimates.AWR requires evaluating logπ H(z|s, g) for the weighted MLE objective.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL ELBO/ODE estimates.AWR requires evaluating logπ H(z|s, g) for the weighted MLE objective

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.857563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:82095843e88d4122429e4d61414c44a0bccf1706b68e79256b15e9c655171193

Observation 3c5906cd-6e4a-4294-b927-78de85192af1 · outbound

This paper cites an unresolved cited work.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-07-10T16:47:24.863650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:aef967b8eed8572907e348d4fe33762869114f9ae41407bce9cedc2775c267d3

Observation 8132e8c2-e28c-4777-85a1-69e8c97f5748 · outbound

This paper cites valid corridor.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL valid corridor

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T16:47:24.817793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:0933fa183b54e267393ce263476c26243502f401449eb7af044d55533f121165

Observation 27527219-e277-442d-8551-91a8c3ef2063 · outbound

This paper cites lucky exploration.

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL lucky exploration

Reference 57

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T16:47:24.794535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-10T16:39:10.354308Z digest=sha256:de46f4b1e8766e40b6539bfb5dfbff7b8b13db9ecdc3588d892d5b2a9d2035d2

Pith citing papers

Observation 8a9a3960-da33-4586-a473-ad6e4c61c817 · inbound

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention cites this paper.

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-02T04:15:52.407616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:15:52.407616Z digest=sha256:5554a3653c369f2f83b47aca2580b5d6a4985bb75df6cd16d024fa5d66585cd7