Pith. sign in

Paper Citation Record · LEDGER

Policy-Based Trajectory Clustering in Offline Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2506.09202.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09202 v2

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:02.986353Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:27:47.896922Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:20:00.517702Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c47bc253-0561-4887-8529-9a3809064937 · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.300603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.962537Z digest=sha256:21abe854336092307934baceed44ee6b20213aabc7339814dda012c9492bb6eb

Observation e1c511b1-b97e-4a24-ac76-3583315dfd99 · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.286204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.966492Z digest=sha256:907099c42540d3ff8ceb1db6db0384ddc0adb98cad8d835b5b5a107541d0ddd4

Observation 379f1434-9361-4d59-8b59-4ccbd6333370 · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.271406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.970261Z digest=sha256:6de71ee532147ed352d7d79647533d2f55c98e374feaccecfc6966f33a01ca13

Observation 86f73a97-a64a-4d6b-9036-cd7c669c123f · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.256975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.974127Z digest=sha256:9b6498b8f7d9a8d6c19688bbab405b6dd44a8038e49337b484009ceaa4be9887

Observation 9c94e70c-2d7b-46da-b108-9da3b240bee5 · outbound

This paper cites Takeball There are four different rule-based policies, i-th policy will pick the i-th ball first.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Takeball There are four different rule-based policies, i-th policy will pick the i-th ball first

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:03.243513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.977885Z digest=sha256:55290d3be8d0ec9e706f419f170ee2d0193ed7dd4a94e880e7fec962cb9a0d13

Observation 2cadc6d3-6a9d-4f48-83b9-e7a07950a58f · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.229272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.981668Z digest=sha256:ab9d3739132302e92af875ccd2bc0a3293e4ff35af60ecf99676b4fce64120e9

Observation adb002fe-3f00-4860-a89e-3e72b48120bc · outbound

This paper cites These policies are corresponding to: No preference, prefer to go up first and prefer to go right first.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning These policies are corresponding to: No preference, prefer to go up first and prefer to go right first

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:03.215458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.986353Z digest=sha256:01e4db2850986bac07aa9639af98a1aaec518ef2f926f0fd50158289c8365a1f

Observation d36b757d-f311-4590-ad1b-b9c86dbe2772 · outbound

This paper cites Unsupervised Deep Embedding for Clustering Analysis.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unsupervised Deep Embedding for Clustering Analysis

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:02.957217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:02.957217Z digest=sha256:612ea0c693806a095eb73c25dd5711cea98373afa2178a946e9d841b6ae852b4

Observation 7d8c3380-d34b-49f4-9707-2f6ca7722d5e · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Behavior Regularized Offline Reinforcement Learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:02.952757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:02.952757Z digest=sha256:e7e56bf66c7bedb189f1f2269275b1ec82e744c90769051ce49e8199652191c1

Observation 1b92c03a-6ad1-47e1-9b08-ad2ba2ea496f · outbound

This paper cites Contrastive Clustering.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Contrastive Clustering

Reference 2020

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:01:03.186412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.937922Z digest=sha256:597c7f8f7dc237243feb69114ee9b2274b446b5fafa2002627946703e0f380df

Observation 2e6a45a0-9bee-4608-80df-30b917f176a3 · outbound

This paper cites Deep Reinforcement Learning for Autonomous Driving: A Survey.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Deep Reinforcement Learning for Autonomous Driving: A Survey

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:02.932706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:02.932706Z digest=sha256:260e86af47bea5f062be745744dd627e672dc7d2827a447f404639c7321aa917

Observation 80ea89f9-9ee0-454e-bc69-aece7f06d8c5 · outbound

This paper cites Dataset Clustering for Improved Offline Policy Learning.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Dataset Clustering for Improved Offline Policy Learning

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:01:03.060042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.947723Z digest=sha256:37315bb60e3f102e21d922d15920f66729de2055267f3ff781356730f5205788

Observation 7d86d253-c3c4-42db-ae9c-ef56f3c9b302 · outbound

This paper cites URL http://dx.doi.org/10.1109/TNNLS.2023.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning URL http://dx.doi.org/10.1109/TNNLS.2023

Reference 2388

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T05:01:03.166577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:01:02.942412Z digest=sha256:e205bd24d374ba0d52080ee3393e79d33abcb18dcf3bb13545457b8c53bef9c0

Pith citing papers

Observation 513d2eda-a545-4fcd-b8e7-42e07391a589 · inbound

Implicit Neural Representations of Individual Behavior cites this paper.

Implicit Neural Representations of Individual Behavior Policy-Based Trajectory Clustering in Offline Reinforcement Learning

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:17:48.435994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T10:27:47.896922Z digest=sha256:d1b00bfb43661926831c7ea000706bda1bf8cc1ef88d1957dabc5398723b7733

Observation 7aa0886d-e7c1-4479-8a86-934233ff7c8d · inbound

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning cites this paper.

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning Policy-Based Trajectory Clustering in Offline Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:20:00.519286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T23:46:50.183064Z digest=sha256:efda893072613e7cf6212e22d94979d333e7fef33fcfd055ec840796efa22fd5