Pith. sign in

Paper Citation Record · LEDGER

RvS: What is Essential for Offline RL via Supervised Learning?

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2112.10751.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2112.10751 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:45.370476Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:57:25.909944Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7dfa7fc1-3245-47dd-9d6d-545b04214591 · inbound

Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL cites this paper.

Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL RvS: What is Essential for Offline RL via Supervised Learning?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:45.370476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:14:45.370476Z digest=sha256:fc08ad1aad56aa915359bb706f5c140ef5dd0d249b5215bd98145f61cb10b474

Observation b26c610a-459b-44af-9acb-69a25dc55670 · inbound

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization cites this paper.

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization RvS: What is Essential for Offline RL via Supervised Learning?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:25.775369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:05:25.775369Z digest=sha256:1f7102869bd7b47264b32209a663f67cd52d0c6b8122c1ee48f25b6257150bb2

Observation 77a78522-f49a-4eb8-96e5-bd86295f3c90 · inbound

Behavioral Exploration: Learning to Explore via In-Context Adaptation cites this paper.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.933445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.933445Z digest=sha256:f624612e24e6c27ffc15c374ab2a91ce573cb2b3ff07c16ddb76db6f19408f87

Observation 97e7a599-1719-4386-aaea-c5b1a5d02a0e · inbound

Hybrid Sequence Modeling and Reinforced Verification for Controllable Target-Conditioned Decision Making cites this paper.

Hybrid Sequence Modeling and Reinforced Verification for Controllable Target-Conditioned Decision Making RvS: What is Essential for Offline RL via Supervised Learning?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T17:23:27.070458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:23:27.070458Z digest=sha256:76151348b59025d8613bfa32d83ae39a8bdf26482f80eda0de0acee70c8b0b7b

Observation abfabb09-a6fb-48e8-8cc1-5b0bbb4ce9c9 · inbound

Generative Sequential Notification Optimization via Multi-Objective Decision Transformers cites this paper.

Generative Sequential Notification Optimization via Multi-Objective Decision Transformers RvS: What is Essential for Offline RL via Supervised Learning?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:57.836508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:57.836508Z digest=sha256:8601ac9ccab6e0d3dd9760ac16f373d85dc64605b9e8a32367398c33b1160177

Observation 7d9d66ef-e6a5-418e-b356-5d4829068f98 · inbound

Nonreciprocal current induced by dissipation in time-reversal symmetric systems cites this paper.

Nonreciprocal current induced by dissipation in time-reversal symmetric systems RvS: What is Essential for Offline RL via Supervised Learning?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T09:47:32.460955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:47:32.460955Z digest=sha256:832b68193963b7769311ede110689320adcb6b0a018dc387cb034190c95b0344

Observation 79207e7f-55cb-4d85-bec5-370e17248870 · inbound

Receding-Horizon Control via Drifting Models cites this paper.

Receding-Horizon Control via Drifting Models RvS: What is Essential for Offline RL via Supervised Learning?

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:35:51.780977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:41:20.941243Z digest=sha256:328d6d3e833e79904ed05c790ee23eec6c9870c1cbc922f344f602a3d75e7493

Observation e164e9d8-980f-4d36-b7ed-737e89913485 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL RvS: What is Essential for Offline RL via Supervised Learning?

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:58.020372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:1f51446789208671e77074dfaa9f4602dc9eac00dbcd214832474acb45d3e621

Observation e0f9405d-fcd0-430b-9810-a5586012827f · inbound

Dash2Sim: Closed-Loop Driving Simulation from in-the-wild Dashcam Videos cites this paper.

Dash2Sim: Closed-Loop Driving Simulation from in-the-wild Dashcam Videos RvS: What is Essential for Offline RL via Supervised Learning?

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:57:10.317901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:15:08.063308Z digest=sha256:39d73d51a2bcc61eabe62b3ebf5a2fca64ff19093bc7e25707e470ed480b05e2

Observation bd9b792f-5df6-4528-b42d-a1c20966e1ec · inbound

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies cites this paper.

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies RvS: What is Essential for Offline RL via Supervised Learning?

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:57:25.911641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T19:23:32.966503Z digest=sha256:68e778d83b01d6d5e7b21b6b79ef56f5e2df0644ce49a99a36de0d451dc40108

Observation fa82bc0e-5d86-4685-ac2b-9c256cdc4859 · inbound

Freeform Preference Learning for Robotic Manipulation cites this paper.

Freeform Preference Learning for Robotic Manipulation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:55:41.940244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T04:58:28.971536Z digest=sha256:8abdaffff1e499dd9eb7768acfd26c64371ecae602d4886ff7636a41c15acb14

Observation 3303363c-3c5c-4159-bde0-22b1bcc557b5 · inbound

Freeform Preference Learning for Robotic Manipulation cites this paper.

Freeform Preference Learning for Robotic Manipulation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T16:55:18.028851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:55:18.028851Z digest=sha256:ab2d67f53e71b41b939ac767bc283ef3f8aeeb155911ab566d39e658ed7209f7

Observation bc371d39-e35f-4d23-a2d4-32ee4f41e98d · inbound

Freeform Preference Learning for Robotic Manipulation cites this paper.

Freeform Preference Learning for Robotic Manipulation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T02:32:20.993278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:32:20.993278Z digest=sha256:efe2421ae8dc9e56e15227d749fc39ae5bcdb099635f2e482b67c9b947608174

Observation 9adc1b62-d588-4645-b977-26d4ae74af3a · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models RvS: What is Essential for Offline RL via Supervised Learning?

Reference 173

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:14.035207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:14.035207Z digest=sha256:42b43743db1bd72117acc7d8d669381a18a3a890473328881fb8edafe51719e0

Observation e8cb85e9-1b8f-4505-bafe-81e03d552cd5 · inbound

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play cites this paper.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play RvS: What is Essential for Offline RL via Supervised Learning?

Reference 1978

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:10.784085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:10.784085Z digest=sha256:a39bda11b297f80e15f794fe9bbea967f7ae900a3becadc308219afd3653669e