Pith. sign in

Paper Citation Record · LEDGER

CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:1902.05605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1902.05605 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:12:08.391612Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T12:35:49.046229Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 148ed036-5575-478d-9161-dbfcf92f9950 · inbound

Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog cites this paper.

Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T12:35:49.049370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T12:32:09.940698Z digest=sha256:dff78af78ddd43cd28b45bc6aa63558f3fbc9246ac05e45a192a86cedb8a0f21

Observation 26523690-016d-4f4b-9dc7-d282437ea02a · inbound

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning cites this paper.

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T19:12:08.391612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:12:08.391612Z digest=sha256:5edb4ad62498e5d43ca3f28eb2a7a44264dcfa5c22693f0202bdbdda28e44037

Observation 2a1b755e-7fb3-4ed6-92a5-4092499cd511 · inbound

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control cites this paper.

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:20:00.681688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T12:17:59.392158Z digest=sha256:1c4ef8462a6788c6c2ee3369b7358b5d0b561e16d397a52a3d90a93090d0cbed

Observation 7108ded1-49ab-4188-a104-ff0323b720a5 · inbound

Distributional Value Estimation Without Target Networks for Robust Quality-Diversity cites this paper.

Distributional Value Estimation Without Target Networks for Robust Quality-Diversity CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:14:46.868265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T00:11:04.222842Z digest=sha256:1ef1d90c34dead635079e03efb857a92960d21e43736a4314efe367056f0c698

Observation 44eba5d9-dc3b-4dfd-9133-5f4843611329 · inbound

AdamO: A Collapse-Suppressed Optimizer for Offline RL cites this paper.

AdamO: A Collapse-Suppressed Optimizer for Offline RL CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:31:03.438730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T14:48:05.166453Z digest=sha256:8c1c56f75c197084704f3af28d85cfbafeac152beb5ab14b686bdf4bc083d0c9

Observation a9aed3e1-703b-4f9f-b7f4-13b961551d02 · inbound

Relative Value Learning cites this paper.

Relative Value Learning CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T08:32:00.068959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:32:00.068959Z digest=sha256:03f5d297d9dfd1c1d3b085e65467c5aadd728ae8876a1b9248e589d406b2054f