Pith. sign in

Paper Citation Record · LEDGER

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning

As of 11 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2502.04141.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04141 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:27:10.256654Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a0ee6b98-8d19-493a-acb8-983a37f0329f · outbound

This paper cites an unresolved cited work.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-08T23:27:10.320069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.251510Z digest=sha256:61846cc03c2ab480996183917c9c1f275a239589d3a527ff814d95e66ede5241

Observation 3371d497-453e-47b3-9549-3da020429228 · outbound

This paper cites Don’t change the algorithm, change the data: Exploratory data for offline reinforcement learning.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Don’t change the algorithm, change the data: Exploratory data for offline reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.357284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.238637Z digest=sha256:c9bcfcb98d49f57b027f62410567cd692f8fe11649bdf3656c8413be0641d12d

Observation 4e31b0ac-45c7-44f0-b023-3c3d684692b6 · outbound

This paper cites an unresolved cited work.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-08T23:27:10.342980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.243975Z digest=sha256:9bd22a5c77f804f9b99ac6c61794c6513d37359e3acaa50b47177c5f635b0d3d

Observation 5fb56640-525d-485e-b339-34dcd82e6144 · outbound

This paper cites For a given set S ⊂ X, radius r, and m > 0, let N (S, r) denote the covering number, the minimum number of balls of radius r needed to cover S.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning For a given set S ⊂ X, radius r, and m > 0, let N (S, r) denote the covering number, the minimum number of balls of radius r needed to cover S

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.327743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.249107Z digest=sha256:8fd42c08315c0e3123bb23b96404c61e6346a5f98905c04dff162d78c0550339

Observation 3d38dfc5-2e18-4919-921c-bcbf517335cd · outbound

This paper cites First notice that E h bH B,w k,n (f ) i − H B,w (f ) ≤ E h bH B,w k,n (f ) − H B,w n (f ) i + E H B,w n (f ) − H B,w (f ).

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning First notice that E h bH B,w k,n (f ) i − H B,w (f ) ≤ E h bH B,w k,n (f ) − H B,w n (f ) i + E H B,w n (f ) − H B,w (f )

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.312568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.254021Z digest=sha256:2d89309f5b2c7912c2fd0c543a487183d9ee04185f80ed64cee45c6b6e361fd6

Observation b2335f68-b416-41a2-b87a-157b2b308bfc · outbound

This paper cites Lemma 1 ((Singh & P ´oczos, 2016)).

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Lemma 1 ((Singh & P ´oczos, 2016))

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.335220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.246550Z digest=sha256:a846f4576669daade6b154e36ed4ddddc36aac063a1dd17ba2764c69d4d287f8

Observation 90271f5d-2c22-4c4b-9614-9d03ead30a03 · outbound

This paper cites Initial trials showed q ∈ {2.0, 3.0, 5.0} led to performance no better (and usually worse) than q = 1 .1, so offline RL training for these q values was not performed.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Initial trials showed q ∈ {2.0, 3.0, 5.0} led to performance no better (and usually worse) than q = 1 .1, so offline RL training for these q values was not performed

Reference 512

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.304572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.256654Z digest=sha256:862bc5f02cd03a559147428faabbfa3257da0bd233fb963a1e56f16a88b4b5db

Observation 59174e23-fbc8-424d-88be-371f12472bc7 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.232623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.232623Z digest=sha256:436107ca658373c3e7f2bac68f4e8ac4c560b2cb2bb5ae804e1d69a6d31d8c70

Observation cb4e6656-b5b8-440d-9028-9478df7a28fb · outbound

This paper cites Diversity is All You Need: Learning Skills without a Reward Function.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Diversity is All You Need: Learning Skills without a Reward Function

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.225010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.225010Z digest=sha256:1bc545008c5195056f60f14f9ef3e8d5e2cbdf68d5e102e4b08349d331c5ce5d

Observation aef9f8ea-4b48-4488-a035-d6eb17167fa8 · outbound

This paper cites nearest neighbor.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning nearest neighbor

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.369702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.222071Z digest=sha256:396c123a21aac5eb577d0f6a68a1afc75fd4a2c673afdcc7fc7798f77e05e1f0

Observation ce4c5cfd-5973-4a25-9d53-cb2b11805480 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Curiosity-driven exploration by self-supervised prediction

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.363276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.235849Z digest=sha256:e7df21fffcfd4469b1530ef7407298da41d8a318c0aae310debba1acca43775a

Observation 8ec95f17-9622-4bb3-9f8f-e81737bb8dd6 · outbound

This paper cites URLB: Unsupervised Reinforcement Learning Benchmark.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning URLB: Unsupervised Reinforcement Learning Benchmark

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.227715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.227715Z digest=sha256:f671cf3498cebfddd06776c65bd653638b4b5a4334c27a258e1609c193525bf8

Observation 73e7b769-2705-4227-8f70-f0cb7359addf · outbound

This paper cites Efficient Exploration via State Marginal Matching.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Efficient Exploration via State Marginal Matching

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-08T23:27:10.230097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:27:10.230097Z digest=sha256:9ec4cb490090880fef32f24ec1964e4970de47a256d4bebfd509dea4bb41e669

Observation 3e23412f-19d2-49a1-b60f-83d48435eb76 · outbound

This paper cites Fix a p.d.f.

Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning Fix a p.d.f

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:27:10.350528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-08T23:27:10.241185Z digest=sha256:42d89a6fe712e622778cb66be4ef7d7eb68722edd6b29eb38cf7bc30437d631b

Pith citing papers

No inbound Pith citation observations are available.