Pith. sign in

Paper Citation Record · LEDGER

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning

As of 6 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2605.01663.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.01663 v2

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact3
  • verified fuzzy5
  • unresolved4
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 06592195-f886-47aa-9f9d-aa729670e0ec · outbound

This paper cites Offline Reinforcement Learning with Universal Horizon Models.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Offline Reinforcement Learning with Universal Horizon Models

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:45:11.864147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:80b2713bfcce143c854fa798d720ad197848695c6c1238d712af4bc5d3736122

Observation 46f30a7e-6ae1-47f7-a45d-6df236bff319 · outbound

This paper cites Engel, Y ., Mannor, S., and Meir, R.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Engel, Y ., Mannor, S., and Meir, R

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T08:13:30.834562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:9c2dac74b0feecb48ac7998769fb7733a70cf3a760d5817de4b2ddb1298b82ab

Observation 0650a644-d299-4352-8cf1-3798e06b25ca · outbound

This paper cites doi: 10.1145/1102351.1102377.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning doi: 10.1145/1102351.1102377

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:45:11.406384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:d3bee953468feb6bb65d8300398c25c690c0451a174c72b8359bf9b1ab64c198

Observation 7fdfa13b-b218-4768-b952-242a14399073 · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:45:11.861707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:94caeb91f813ca43407c956485c281da821fbff8091a809cfe4494c38c5cd393

Observation b817a577-db02-4820-96c6-d17b24ff56f8 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:45:11.856744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:e3fa734a7784b0537b7944a57d872a8c7080dd78635c2d5caf03a4fde59b8ef7

Observation f8c24e63-fa9f-4c11-a987-a079e5d245f6 · outbound

This paper cites Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:45:11.859295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:fcf8e03a7a6512283362589a75dcfb9d64f92150a03360341a0db37911666210

Observation c699f7bc-d823-4e95-9f11-996bd42ae390 · outbound

This paper cites Kostrikov, I., Nair, A., and Levine, S.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Kostrikov, I., Nair, A., and Levine, S

Reference 7

Resolution
malformed identifier
raw_fallback, observed 2026-07-07T08:13:30.836113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:5bfd0cef83d65df8c97757e191ad3db7280a2a40caab0d1f91a9fa377d8f9772

Observation 8870bb17-d89e-4d6d-8ab0-8a1db1903b7d · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-01T00:45:11.403899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:95e76fe83f149d1cabe586a762c16503b89afb05f1a40144157319217ba26c40

Observation 13f7a8d5-8de7-4e34-bf00-b48941cc9659 · outbound

This paper cites Nguimatsia Tiofack, F., Le Hellard, T., Schramm, F., Perrin- Gilbert, N., and Carpentier, J.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Nguimatsia Tiofack, F., Le Hellard, T., Schramm, F., Perrin- Gilbert, N., and Carpentier, J

Reference 9

Resolution
verified exact
doi, observed 2026-07-01T00:45:11.408358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:f3a886169e996438bcf921e9f17687d40fef1c27c4038c3be48c48ab0b7da468

Observation b03f38db-1a2f-4d7e-a7f7-5f1aba189b51 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:45:11.854355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:59d49a97d4c3c82a5fd48c8ee98ed3f908f44d2b27fc86a2ddc0456ba213621d

Observation 87231a91-efab-4702-a5d1-c858239dfe4a · outbound

This paper cites Srikant, R.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Srikant, R

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T08:13:30.837883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:d4d9664351cbf71ffbf0b1ba41e284d1ed958ff77b254e8978867832b694d4fb

Observation 33c5469a-74d7-454f-aa19-d8aca2b0eb94 · outbound

This paper cites Venkatraman, S., Khaitan, S., Akella, R.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Venkatraman, S., Khaitan, S., Akella, R

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T08:13:30.828325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:54c8490ab95b638fe00598a52a8a4ec10cb2407c0ca0c7423aff231607e6419d

Observation 8f3af840-3e14-47c8-becb-804337b8e558 · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Behavior Regularized Offline Reinforcement Learning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:45:11.401553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:b33b6ea2d4842284e04b0496709af3cb6a6c196a08bdb56657ff20fde64466f3

Observation 88b1737f-5b0a-4d65-a271-04ac47710991 · outbound

This paper cites an unresolved cited work.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-07-07T08:13:30.829916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:00874019fd2d0b567d2192cac2e50f8008cc8b1a71dcce5b9744674a2ebe89b5

Observation 4b4afea7-4f81-4996-840a-89195b24e8a1 · outbound

This paper cites Elements of a σ-algebra are calledmeasurable setsorevents.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Elements of a σ-algebra are calledmeasurable setsorevents

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T08:13:30.826630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:8d59f2a51d7780da3c44f9f86effbeed20a8b67fe2200c849f29d5a608e4b5c8

Observation e3baff4b-e64d-40b7-979c-f517f2e85763 · outbound

This paper cites an unresolved cited work.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-07T08:13:30.831451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:5c1d63516d9b6f0968c55a1cbaf67c9c42d744c7e85cd6b5a9f912b934fad933

Observation 7fd3493d-bf7a-4ee8-9230-dfe54b6f53d2 · outbound

This paper cites The Adroit tasks require learning complex skills such as spinning a pen, opening a door, relocating a ball, and using a hammer to hit a button.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning The Adroit tasks require learning complex skills such as spinning a pen, opening a door, relocating a ball, and using a hammer to hit a button

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-07T08:13:30.824737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:668d0c90b706bd2bb52b00d1a8028c1155130fec017e045bc967adba13e84c6c

Observation b1235a75-fecf-4497-bbd5-b9d0121f89fa · outbound

This paper cites an unresolved cited work.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-07-07T08:13:30.822985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:e10fbbdafbd592eb6488c07d227aee6ec97181cbc9e685de84c0a1706af192b3

Observation bd8ec583-f733-49bc-b032-b0fd35405e7e · outbound

This paper cites an unresolved cited work.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-07-07T08:13:30.833038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T00:43:24.171887Z digest=sha256:640c4216db1d5b3d07f67242062826decfa88eeb980b4ed0b7e75cf5f5007679

Pith citing papers

No inbound Pith citation observations are available.