Pith. sign in

Paper Citation Record · LEDGER

Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2211.15144.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.15144 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T08:05:47.128354Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:15:32.316103Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e4bbf9d-5b3a-4123-8bbd-95e07e88f81b · inbound

Learning Interactive Real-World Simulators cites this paper.

Learning Interactive Real-World Simulators Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

Reference 258

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:15:18.597937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-16T02:15:18.265190Z digest=sha256:72611923e16fad78ed5699aee8389f547bb75f8ee9ad882130711a8c9940cdc7

Observation 58facab5-fe10-48e6-b4a2-c0a0c40ca365 · inbound

Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning cites this paper.

Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:21:13.579804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:19:28.727598Z digest=sha256:08efbb138428da97300da4210e259ec541d8032adc59f3cbd3a294504c853330

Observation 547a2e29-6d8e-4bfc-842b-cd9210aa0923 · inbound

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies cites this paper.

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:36:12.509676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T19:31:38.069592Z digest=sha256:531a80d5b036a3a9d3ecbc6ac8069456b5293e9f656068fd65272ed9981f37ab

Observation 2c291311-8889-4673-b059-44ecf5f1b6c6 · inbound

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies cites this paper.

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:15:32.319056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T08:05:47.128354Z digest=sha256:d9eed539576607d0c2811c8e7d9db27560a4ecfc8c16e39516346532af1bb229

Observation f633221a-fd9a-407d-8c77-5a26d9c64a93 · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:56:00.648608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T16:25:25.739019Z digest=sha256:96f33236a2276eb44a784eb52955ce5dd9990f266f66ffe9d77a3e79e8c288cf

Observation 43bb3947-5eac-42dd-8b84-b09c3f92b383 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes

Reference 81

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:30:58.205675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:c04a46bc2e3d021a6e12bc9c8516f94b7629947fc3cb42b4dde52cdb96782be1