Pith. sign in

Paper Citation Record · LEDGER

OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2010.13611.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.13611 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:00.977525Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:57:29.181327Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e67ef3e2-5dc7-4102-b3b9-af35d619c232 · inbound

Decision Transformer: Reinforcement Learning via Sequence Modeling cites this paper.

Decision Transformer: Reinforcement Learning via Sequence Modeling OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T15:11:11.314700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T15:11:11.056013Z digest=sha256:b7d11b5951dfc19bafc6361f8f679af0098b2fbf6b6f071b4c007029724ffba0

Observation 24ac5f17-dd9e-4fd0-82fd-598970556c8f · inbound

What Matters in Learning from Offline Human Demonstrations for Robot Manipulation cites this paper.

What Matters in Learning from Offline Human Demonstrations for Robot Manipulation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-13T08:51:55.958219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T08:51:55.826747Z digest=sha256:68695ae6f670176248683bdd5f57b4b2382be9171f2351e20f2e2f1518b8f78f

Observation 0d6edf85-53ae-4327-bf2b-0235d61acc91 · inbound

Goal-Conditioned Supervised Learning for Multi-Objective Recommendation cites this paper.

Goal-Conditioned Supervised Learning for Multi-Objective Recommendation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T17:32:06.994818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:32:06.994818Z digest=sha256:f11c0c2fc56908e08dd90c0484b3f07121825f68e482bc5eb66669747467329b

Observation e7a7730c-45b8-4919-b357-713eb95f3b6b · inbound

GROOT-2: Weakly Supervised Multi-Modal Instruction Following Agents cites this paper.

GROOT-2: Weakly Supervised Multi-Modal Instruction Following Agents OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:41:44.970264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:41:44.970264Z digest=sha256:2818174ab6d3083d6c341a5cbcec5fc3e660967b09be032f13b4e25a33b23771

Observation 9ebb9f36-1bc3-48b6-bb20-4d7db3969973 · inbound

NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations cites this paper.

NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-10T17:00:39.289737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:00:39.289737Z digest=sha256:816fd46c2da28f6aa17e3698dc67d6add6781bbbf317abf4379a0345487f2e0b

Observation 10d5154d-f5f0-4b85-98ca-1fa9959b4078 · inbound

Dynamic Contrastive Skill Learning with State-Transition Based Skill Clustering and Dynamic Length Adjustment cites this paper.

Dynamic Contrastive Skill Learning with State-Transition Based Skill Clustering and Dynamic Length Adjustment OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:00.977525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:46:00.977525Z digest=sha256:62abb4efb75048f19491f61eb124c004954d43ffc7dc4f84fcaba1c38b3aeb05

Observation ff9dbc7f-779e-41fd-89b2-78b7f97a0a6c · inbound

Behavioral Exploration: Learning to Explore via In-Context Adaptation cites this paper.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.139432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.139432Z digest=sha256:9c56944f679f50ec7866eb801e0d7b68afc98116b200ac8cb34f3b6753793a97

Observation 1192f49f-ab8d-4644-805a-7744d51c4ed3 · inbound

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs cites this paper.

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:34.847846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:17:34.847846Z digest=sha256:2ba715ac17d701673caa75f6669f7e7f7e23d7b44c804c9a252fa08f015acc57

Observation fbbd8599-f619-4da0-873b-2e186e9b76db · inbound

Learning Upper Lower Value Envelopes to Shape Online RL: A Principled Approach cites this paper.

Learning Upper Lower Value Envelopes to Shape Online RL: A Principled Approach OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T08:45:47.978112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:45:47.978112Z digest=sha256:2d42852453d07604bce28d2dda6a67a608b6560410e0495f38e1dcb205594c20

Observation 384aca80-181a-45a0-93e7-fffcd6613991 · inbound

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation cites this paper.

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T15:05:01.788849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:05:01.788849Z digest=sha256:52c4b2d829c4e0a608bd676c1ea71952e761c04a19e85492672b4c437442beb9

Observation 22e07cd7-df74-4651-be1c-1af7f93bbba8 · inbound

Implicit Safety Alignment from Crowd Preferences cites this paper.

Implicit Safety Alignment from Crowd Preferences OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:31:16.907176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-22T08:28:41.652865Z digest=sha256:c29a0adf7f38b94de8a2b88ce7d2f3a24984031067be7849788bcc66b260d6fa

Observation 1fa937b6-3149-4cd3-89a4-eddf17901ccd · inbound

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement cites this paper.

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:57:29.182689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T17:33:35.857240Z digest=sha256:6b5a271985d95a8050c7e35dadc21f47785db0fd80c4e81dffdfad8e97bdfbe6

Observation 097a2308-2c05-487d-91ef-4ae79db223c1 · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.212408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:07c4688e23839430b5a8c90b4f864041938a94c8480c4ddc16adc3aa0950cbd2

Observation 01c0d124-f73a-4f69-979b-8e84014c9005 · inbound

Adapting Generalist Robot Policies with Semantic Reinforcement Learning cites this paper.

Adapting Generalist Robot Policies with Semantic Reinforcement Learning OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:45:42.729562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T05:09:29.625066Z digest=sha256:ab8d6a2a172d783d7d4bee6ddb52c2593ad2f5c503b40926237f5bc38d1bb39e

Observation ed133a94-6119-455a-9f1d-3b4e59acae50 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:26.462169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:26.462169Z digest=sha256:61dd0c96aac0d78063e553a9bfd811408855601ee60e0243b15796061625e8b8

Observation 31556764-0c0c-4591-a04c-39bbbfdc47b2 · inbound

Convex-Hull-Neighborhood Smooth Dual Generalization: Controlling Local Correction Propagation in Offline RL cites this paper.

Convex-Hull-Neighborhood Smooth Dual Generalization: Controlling Local Correction Propagation in Offline RL OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:57:30.167094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:57:30.167094Z digest=sha256:56139ffa0fe2628b977e2082e7e65030c3813e0439756cd904cb273ae2880bc2