Pith. sign in

Paper Citation Record · LEDGER

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2502.04141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04141 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:27:10.256654Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a0ee6b98-8d19-493a-acb8-983a37f0329f · outbound

This paper cites an unresolved cited work.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-08T23:27:10.320069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.251510Z digest=sha256:90e28a2905e53989750098942ec34639afedefb24fcc27a4cfa1412c4a3a6315

Observation 3371d497-453e-47b3-9549-3da020429228 · outbound

This paper cites Don’t change the algorithm, change the data: Exploratory data for offline reinforcement learning.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Don’t change the algorithm, change the data: Exploratory data for offline reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.357284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.238637Z digest=sha256:343b8adc5fc3e63e67e631ddb720cb50b674fb13f55ddafd515ef29b135731e7

Observation 4e31b0ac-45c7-44f0-b023-3c3d684692b6 · outbound

This paper cites an unresolved cited work.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-08T23:27:10.342980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.243975Z digest=sha256:482e68606496f52829d52a6b04672196510a43ca06b78239c8ed83221d0a66ce

Observation 5fb56640-525d-485e-b339-34dcd82e6144 · outbound

This paper cites For a given set S ⊂ X, radius r, and m > 0, let N (S, r) denote the covering number, the minimum number of balls of radius r needed to cover S.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning For a given set S ⊂ X, radius r, and m > 0, let N (S, r) denote the covering number, the minimum number of balls of radius r needed to cover S

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.327743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.249107Z digest=sha256:5dbb0d7126944f0602160bea9b61be637f1e28e67276a73f6a1e9ffc899ffc7d

Observation 3d38dfc5-2e18-4919-921c-bcbf517335cd · outbound

This paper cites First notice that E h bH B,w k,n (f ) i − H B,w (f ) ≤ E h bH B,w k,n (f ) − H B,w n (f ) i + E H B,w n (f ) − H B,w (f ).

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning First notice that E h bH B,w k,n (f ) i − H B,w (f ) ≤ E h bH B,w k,n (f ) − H B,w n (f ) i + E H B,w n (f ) − H B,w (f )

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.312568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.254021Z digest=sha256:7a9e011f37399739731d6d80e29b455cf75c8e73a3d3f0c7d3f39cc1e4e461a5

Observation b2335f68-b416-41a2-b87a-157b2b308bfc · outbound

This paper cites Lemma 1 ((Singh & P ´oczos, 2016)).

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Lemma 1 ((Singh & P ´oczos, 2016))

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.335220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.246550Z digest=sha256:b8f5e899fe77cd76ed1e6b469665abceae71d225cc7a6506ec8a5c423606d56f

Observation 90271f5d-2c22-4c4b-9614-9d03ead30a03 · outbound

This paper cites Initial trials showed q ∈ {2.0, 3.0, 5.0} led to performance no better (and usually worse) than q = 1 .1, so offline RL training for these q values was not performed.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Initial trials showed q ∈ {2.0, 3.0, 5.0} led to performance no better (and usually worse) than q = 1 .1, so offline RL training for these q values was not performed

Reference 512

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.304572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.256654Z digest=sha256:e9e3619ab89f9e490728c640c43d65b2d68be367505eb7907567f4baa7a15c44

Observation 59174e23-fbc8-424d-88be-371f12472bc7 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.232623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.232623Z digest=sha256:436107ca658373c3e7f2bac68f4e8ac4c560b2cb2bb5ae804e1d69a6d31d8c70

Observation cb4e6656-b5b8-440d-9028-9478df7a28fb · outbound

This paper cites Diversity is All You Need: Learning Skills without a Reward Function.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Diversity is All You Need: Learning Skills without a Reward Function

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.225010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.225010Z digest=sha256:42c8a8be3237f898692dd3de8bb3c7ce9ed4a0a910902d679171235594ea1f0a

Observation aef9f8ea-4b48-4488-a035-d6eb17167fa8 · outbound

This paper cites nearest neighbor.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning nearest neighbor

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.369702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.222071Z digest=sha256:1c16ebc7611d3754d4c5bb7f45912b9cb3be20a8e0be4c00fbcee6726b1c71dc

Observation ce4c5cfd-5973-4a25-9d53-cb2b11805480 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Curiosity-driven exploration by self-supervised prediction

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.363276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.235849Z digest=sha256:4fa3f4c07b1af1fb3830a2a450e49c539d632a236bca2f29c941efd11aa16213

Observation 8ec95f17-9622-4bb3-9f8f-e81737bb8dd6 · outbound

This paper cites URLB: Unsupervised Reinforcement Learning Benchmark.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning URLB: Unsupervised Reinforcement Learning Benchmark

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.227715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.227715Z digest=sha256:b08ca42dbbf835e65fefeb4b003e0a8d6cbe672ab0c03c69c2ea40ce5e938abc

Observation 73e7b769-2705-4227-8f70-f0cb7359addf · outbound

This paper cites Efficient Exploration via State Marginal Matching.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Efficient Exploration via State Marginal Matching

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.230097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.230097Z digest=sha256:4352a417638dbe31881114c2c1c3a384db4d78c6c3d19cd5dab1881650edec87

Observation 3e23412f-19d2-49a1-b60f-83d48435eb76 · outbound

This paper cites Fix a p.d.f.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Fix a p.d.f

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.350528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-08T23:27:10.241185Z digest=sha256:0f26edb3dd66645a1e0a4e7fe21abafda25874c0d0a87b6783541becc3753edb

Pith citing papers

No inbound Pith citation observations are available.