Pith. sign in

Paper Citation Record · LEDGER

Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:1910.04295.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1910.04295 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:25:50.732625Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T14:09:53.348942Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eec5f744-7b5d-4f39-9d94-85ca914afc08 · inbound

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games cites this paper.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.732625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.732625Z digest=sha256:d368158791ec313406f2fbddf89a1c5e101b7dff6a1d2f3ce4045d87da0991a3

Observation 1157fc72-6af4-465d-b37b-4467c5d41ba4 · inbound

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon cites this paper.

Toward Optimal Statistical Inference in Noisy Linear Quadratic Reinforcement Learning over a Finite Horizon Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:35:24.252521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:35:24.252521Z digest=sha256:0a5aa8a91a0a67e4b802d7500d5fa4bb2b4ae335fd74b25758892cd93a8cd0bd

Observation 48c6c7de-978a-48f0-b6ba-e4f829410715 · inbound

Policy Gradient for Continuous-Time Mean-Field Control cites this paper.

Policy Gradient for Continuous-Time Mean-Field Control Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.507409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f38a726d102373e70374097c336ecc9a2c1c4dad26fcfd5f492837bb8239ae09

Observation aac72c2c-a677-41f9-a90b-50cf2c77c3d6 · inbound

Mean-Field PhiBE: Continuous-Time Mean-Field Reinforcement Learning from Discrete-Time Data cites this paper.

Mean-Field PhiBE: Continuous-Time Mean-Field Reinforcement Learning from Discrete-Time Data Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:09:53.351411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T04:28:37.809929Z digest=sha256:6b8ef58d7ba53dc51599cf4039604dfe030b24433866b731890bfaa8e9379d7c