Pith. sign in

Paper Citation Record · LEDGER

When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2204.05618.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.05618 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:17:38.623944Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

17
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 77463182-120e-4f61-8947-c6eeeacd2db4 · inbound

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training cites this paper.

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:42:52.669484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T04:42:52.627166Z digest=sha256:82f006dbb3350f47c7d18286dbca2f1cf179d0e6b7bb27667b94c90abf41ef5e

Observation 0f15e62c-e660-4d6f-b89c-22f7e4860177 · inbound

Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance cites this paper.

Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T04:52:59.457090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:52:59.457090Z digest=sha256:1c3d59fd4f74a1be0f4f07c0934c757b3f8666c58c1bb6d8c207249aa720df28

Observation c5f9e2b2-8542-4e3f-8f5a-58b6cdc60273 · inbound

An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning cites this paper.

An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T12:17:38.623944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:17:38.623944Z digest=sha256:259a280f8eb4c500b263076fc88879cf61a58b3f841b0a8c3b9c6ec1980d361e

Observation 0079a857-2ee3-49f1-b19c-ebff7a8fdfb9 · inbound

Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference Optimization cites this paper.

Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference Optimization When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:32:52.276927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:32:52.276927Z digest=sha256:e5827de4db5a6b167345c7bfd0c6e22fdc477a4ca7b0940334c07d8e5cfb5c4d

Observation 251fbc9a-589b-4c48-a248-b64ea91a439d · inbound

CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding cites this paper.

CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:03:56.797614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:03:56.797614Z digest=sha256:b1715066c2a4c5ddb11a8ad5cbd8cd5f25bdacb53724545e8216182cf3da9b78

Observation 4402ccd8-964b-4536-8a3f-ec8efab2906c · inbound

Comparing Behavioural Cloning and Reinforcement Learning for Spacecraft Guidance and Control Networks cites this paper.

Comparing Behavioural Cloning and Reinforcement Learning for Spacecraft Guidance and Control Networks When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T15:19:38.190553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:19:38.190553Z digest=sha256:a4cd5783ef4c495717ddcd107a6bd44cf02b98b9bb4ba9c88f9ba276b8a21ea4

Observation e9088da3-274b-425a-af98-145047a21fc7 · inbound

Active Query Selection for Crowd-Based Reinforcement Learning cites this paper.

Active Query Selection for Crowd-Based Reinforcement Learning When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T16:02:25.972654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:02:25.972654Z digest=sha256:51a6bdfb59e8cba7f7a8d460a7ed11d00e2e357c1e8cca5cc7fd89a717c00709

Observation 1a9f12ff-d8f8-4824-bcd6-c7886feacac1 · inbound

The hidden risks of temporal resampling in clinical reinforcement learning cites this paper.

The hidden risks of temporal resampling in clinical reinforcement learning When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:07:29.229254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T07:05:16.379996Z digest=sha256:b3cf961d236cf020918296c45c0c1c7daeb4da83489b815f621ed3fb79b1ddc8

Observation 6dac577b-bd7e-4190-b731-bf7d2642e43e · inbound

An Introduction to Causal Reinforcement Learning cites this paper.

An Introduction to Causal Reinforcement Learning When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 145

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:49:57.169851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T00:17:57.091481Z digest=sha256:cc1db0a3808cb620075ab3113edbbf7ad222c32193ae57cb22403328c8242ad9

Observation 1d99f2f9-11d2-4c26-b1aa-341cf67d56cc · inbound

From Bootstrapping to Sequence Modeling: A Unified Generative Framework for Personalized Landing-Page Modeling cites this paper.

From Bootstrapping to Sequence Modeling: A Unified Generative Framework for Personalized Landing-Page Modeling When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:55:51.391002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T02:56:52.303152Z digest=sha256:f5fb77f43b506d6f356dc19c2e8b1aecb462bf07eca531437d3d471322591297

Observation ffaae6f9-1c0e-43c1-97f3-c6c55e8d6f17 · inbound

ReBRAC-v2: The Return of the King cites this paper.

ReBRAC-v2: The Return of the King When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T00:32:42.895141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:32:42.895141Z digest=sha256:2e94410a3bd2cd01bc67857ae8737e10c15f967569ff54a31d8325eefc6da5b7

Observation 944f50eb-2561-49dc-a84b-812052f33c30 · inbound

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents cites this paper.

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:13:31.897892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:13:31.897892Z digest=sha256:d5f9b970e2c11134920dcdf9b49c919b33e791efa3d09e03962075c190d4b468