Pith. sign in

Paper Citation Record · LEDGER

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2512.18763.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.18763 v2

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T15:00:44.145471Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-30T21:23:27.894346Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved54
  • parse uncertain0
  • malformed identifier4
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cccccf50-b56f-46a0-a4af-9b92bbb9f1e5 · outbound

This paper cites Bertsekas,Reinforcement Learning and Optimal Control.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Bertsekas,Reinforcement Learning and Optimal Control

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:37.980651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:37.980651Z digest=sha256:8814f5f7f4346198a56343a0dc0d432c7e2ac75e4363cf4a6f30511dfeb32b2e

Observation dd8b8b4b-c79d-4b77-9c8d-62f8a1846ec4 · outbound

This paper cites an unresolved cited work.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.103352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.103352Z digest=sha256:61f74397e6ae31d04428b66384d71dfe06fe53ee324d05123c23f73a7ebf284c

Observation 45e33159-f696-47b5-8a06-47aee17f0c0c · outbound

This paper cites Q-learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Q-learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.162131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.162131Z digest=sha256:b9658c71999aed567f21095ce9509b3b682b4bef5d64b6b006848bbbd17e89cc

Observation 538eac2f-24f4-4606-bc6e-41ac1b2b4d77 · outbound

This paper cites Convergence results for single-step on-policy reinforcement-learning algorithms,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Convergence results for single-step on-policy reinforcement-learning algorithms,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.229895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.229895Z digest=sha256:d48d882707ed4c88cbcc9ce6d1ccbbf7e5baeb5c3078b34428b7b60d8b9fb495

Observation 1923cd9a-ca93-42dd-8144-06fcdc64dcbe · outbound

This paper cites Kernel-based reinforcement learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Kernel-based reinforcement learning,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.299601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.299601Z digest=sha256:3e0234f04292b82ddc4fcf4950185dcf43cb1e1e295ba81c871ec0eecd321be9

Observation 3e1b635c-5caa-4c83-af7e-65dabe41ad6d · outbound

This paper cites Kernel-based reinforcement learning in average-cost problems,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Kernel-based reinforcement learning in average-cost problems,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.389699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.389699Z digest=sha256:84ecfd754d37944c1070bc664932e0d37940c4b229e1916ceb826065d846fa35

Observation 6085959e-d528-487c-8c1f-b659e5943ad8 · outbound

This paper cites Stochastic kernel temporal difference for reinforcement learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Stochastic kernel temporal difference for reinforcement learning,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.457558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.457558Z digest=sha256:c60c47da06da728e721bbd14151859982a4a1cc4a04d66855ff128f7619586e9

Observation 800e07e6-e3f2-4c7a-8bb0-b8602b0f40a4 · outbound

This paper cites Learning to predict by the methods of temporal differ- ences,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Learning to predict by the methods of temporal differ- ences,

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-03T15:00:38.540910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.540910Z digest=sha256:a60e73f00afbdc819e776f3f198ce9fcdfb40e072e2e1fcffec664ed31b46b5c

Observation 4d186437-b2a4-478e-aaa4-f711b997671c · outbound

This paper cites Least-squares policy iteration,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Least-squares policy iteration,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.636891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.636891Z digest=sha256:47f1de7f892016198b52fc78ff8392476cbebbda20b4e57340ecfa26179f8b67

Observation 4b082f3d-7235-4fc0-bf83-0d76c3ac3e68 · outbound

This paper cites Regularized policy iteration with nonparametric function spaces,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Regularized policy iteration with nonparametric function spaces,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.701511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.701511Z digest=sha256:bc9b7e8fa727296f77958d584420efe610d37e9d9c40c0714826425e9e1e6d69

Observation 122b5260-d903-4da2-b62a-8355c89fdfc9 · outbound

This paper cites Kernel-based least squares policy iteration for reinforcement learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Kernel-based least squares policy iteration for reinforcement learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.782560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.782560Z digest=sha256:198f40d350d29c36e429e8b502a6dd341997a30b0e49129ef321359613c06d48

Observation e01dbce6-101e-4ce9-8287-40eb7d39c3e7 · outbound

This paper cites Online Bellman residual and temporal difference algorithms with predictive error guarantees,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Online Bellman residual and temporal difference algorithms with predictive error guarantees,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.878074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.878074Z digest=sha256:20b8f669323d4f2b485440ebfc25d9ba4b7df9fe80da19b5a4c6aeb552c8b4a8

Observation aa52abcf-19fd-439a-9829-4646586fa9a6 · outbound

This paper cites Dynamic selection of p- norm in linear adaptive filtering via online kernel-based reinforcement learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Dynamic selection of p- norm in linear adaptive filtering via online kernel-based reinforcement learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.929115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.929115Z digest=sha256:55eaf4dcdcbbe463cc441eea26afd03310b5d88c44af2c8c5d2f885071ad5602

Observation e16d4662-b278-4b2a-9eb6-f5db820748dc · outbound

This paper cites Proximal Bellman mappings for rein- forcement learning and their application to robust adaptive filtering,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Proximal Bellman mappings for rein- forcement learning and their application to robust adaptive filtering,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:38.992047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:38.992047Z digest=sha256:fe08379322ccd134011d59c162c1c2d7fe54afc2c5ad7d1899c38dabd843da24

Observation cbe271ff-fb7e-4a15-ba2e-fe620db329fa · outbound

This paper cites Nonparametric Bellman map- pings for reinforcement learning: Application to robust adaptive filter- ing,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Nonparametric Bellman map- pings for reinforcement learning: Application to robust adaptive filter- ing,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.052473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.052473Z digest=sha256:4b831a87761f72d0824a1d2ad579151efa7a85446d295a54647d3368bc1dabe2

Observation 8355b67a-2076-4a1a-966c-419ac7be4d1f · outbound

This paper cites Theory of reproducing kernels,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Theory of reproducing kernels,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.161258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.161258Z digest=sha256:abf0680568d795b184ea6152c4f2a00e9aebbebc3ce89de61a63f94f292c9b76

Observation 64edac9f-38ac-4f7a-8ed2-2d8dd09b363d · outbound

This paper cites Schölkopf and A.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Schölkopf and A

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.247832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.247832Z digest=sha256:9c3ceb926eca6c4510af635ab063ba23b93f343b38f49a71a090b6e86aea50ae

Observation a1df23fc-7010-44a7-8e20-25959e1f79bc · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.340137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.340137Z digest=sha256:cbde400ad3edf1b8cd8f1823cc890cf6b9d30457f5f3630851fcfd1d1e64fda6

Observation c4ee15a5-c33f-4700-95f4-5bdba6824f16 · outbound

This paper cites Deep reinforcement learning with double Q-learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Deep reinforcement learning with double Q-learning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.427371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.427371Z digest=sha256:cf1f444ec3ce6df9fc086300a4be284c937ce9db186c270a0bc494197a0d8353

Observation d9356b85-9b00-4db6-8eae-84ddc859700d · outbound

This paper cites Reinforcement learning for robots using neural networks,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Reinforcement learning for robots using neural networks,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.491581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.491581Z digest=sha256:318439962105d7cc6b9f0878e95cffcfd960fc442333aa61ccce633f4d6df718

Observation dae98b34-96ee-4f01-978d-cf256a871126 · outbound

This paper cites Reinforcement learning based on on-line EM algorithm,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Reinforcement learning based on on-line EM algorithm,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.633347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.633347Z digest=sha256:a537c05f9baf46597bf5256f1a9e422a268ecd96792ca654913c0bb20ba42b23

Observation f6440d4f-45bc-4ef7-981b-274b00adca48 · outbound

This paper cites Reinforcement learning with Gaussian processes,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Reinforcement learning with Gaussian processes,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.769068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.769068Z digest=sha256:856923774d72704cdacf3aecb4dd3e69be132614d87faf4d53ab38346dd6e1a4

Observation 5d3ac583-5915-4746-96ae-a69f00277375 · outbound

This paper cites Reinforcement learning with a Gaussian mixture model,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Reinforcement learning with a Gaussian mixture model,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:39.932136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:39.932136Z digest=sha256:5d35aa7adfc983b5b7889db603f379893dec01d9560918b07809cc31da5dcec5

Observation bf0cbf36-bf79-4a11-a559-3dab74876e60 · outbound

This paper cites Online reinforcement learning using a probability density estimation,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Online reinforcement learning using a probability density estimation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.039043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.039043Z digest=sha256:e9843e954ca9464756f1b13f902982ed6dc5d57247d27dbbdc4bbe72b1636646

Observation 3a6e18fd-5c8e-48d7-9a55-d43fb261ac90 · outbound

This paper cites Distributional deep reinforcement learning with a mixture of Gaussians,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Distributional deep reinforcement learning with a mixture of Gaussians,

Reference 25

Resolution
malformed identifier
no resolver link, observed 2026-08-03T15:00:40.154676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.154676Z digest=sha256:3a0ffd2383c10c453a22cac3c8e4ad3d0873e510da270db5c5dd4ba7c9223979

Observation dfd0886f-550f-4e81-a439-d071fa951ad9 · outbound

This paper cites Gaussian mixture models,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Gaussian mixture models,

Reference 26

Resolution
malformed identifier
no resolver link, observed 2026-08-03T15:00:40.278616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.278616Z digest=sha256:c845a6b82fe9329fe0a68f72ff3c2ca0742998b0d2bb280b558c25b1166aada4

Observation 2890067a-993b-464f-abfc-6c15bd6fc123 · outbound

This paper cites McLachlan and D.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning McLachlan and D

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.442616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.442616Z digest=sha256:d7063e8ae419f3575ef2caf48d1eb18ab2b1a51c85ceef88295bdbcb215b634a

Observation b75c20cd-f5d4-4155-975e-fe69a9d7b030 · outbound

This paper cites Maximum likelihood from incomplete data via the EM algorithm,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Maximum likelihood from incomplete data via the EM algorithm,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.604180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.604180Z digest=sha256:395ca2b707090e6229777bd339cc948d8828f333ee9c242d1430d5ef83b0c63b

Observation f83b8f9d-cd82-4272-a975-2eb15470fcee · outbound

This paper cites Unsupervised learning of finite mixture models,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Unsupervised learning of finite mixture models,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.682353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.682353Z digest=sha256:2814a41d9b075b31fbebc16cd038fc5b41374fe53ae339369a97cd21855d3067

Observation 33994f8a-b74d-48a4-8b2a-30324f61a1e1 · outbound

This paper cites Absil, R.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Absil, R

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.760731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.760731Z digest=sha256:18148c3711892a02a8031763fb98efdb64553ce9503538d6ae5a2a5ec79ff6d8

Observation 8bce3830-30bc-433a-9a26-5895c20fae36 · outbound

This paper cites Rie- mannian proximal policy optimization,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Rie- mannian proximal policy optimization,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.850046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.850046Z digest=sha256:54383354d4a7b226129edc9526c9918fc4c22178174e81b4711d73edd906b9ed

Observation 61fabfe3-29af-44f3-8604-fb0d6711790c · outbound

This paper cites Policy gradient methods for reinforcement learning with function approximation,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Policy gradient methods for reinforcement learning with function approximation,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:40.914933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:40.914933Z digest=sha256:de0e3457df5ef70f91391e6c829d3217ee0e67dcd95dbb4ef3afe223c4e59240

Observation d87b11ec-d658-4798-9d20-1193cf6bd2d6 · outbound

This paper cites Approximation and radial-basis-function networks,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Approximation and radial-basis-function networks,

Reference 33

Resolution
verified exact
doi, observed 2026-08-03T15:04:08.482603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-03T15:00:40.976534Z digest=sha256:f2686b5f978ad18462a0a2dda006f3c0c5f6694a29b24f08602c59f5b4f3d05f

Observation a34a34e6-1758-4b25-8245-c31e18a4b42b · outbound

This paper cites Riemannian Q-functions for policy iteration in reinforcement learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Riemannian Q-functions for policy iteration in reinforcement learning,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.075599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.075599Z digest=sha256:036029ea0f1c1589ec6eb1cdd62c97da6c3651f40979d47c9da1b095418bd31f

Observation 3ceed7cb-9bdc-4899-b6c0-de888843039d · outbound

This paper cites an unresolved cited work.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.172037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.172037Z digest=sha256:937175970d293745783c4d77ed439974aaf383d0b18afba40df6e79117763a4b

Observation 299d50ca-68bc-4bf7-8b48-5f25f371a6b8 · outbound

This paper cites Tight performance bounds on greedy policies based on imperfect value functions,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Tight performance bounds on greedy policies based on imperfect value functions,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.309276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.309276Z digest=sha256:df7a9aace47778f7cb652270ad582d63d65feee730ac9f937fadd843e4bbfe4f

Observation d6a478d8-e42e-4032-9dfd-7eb9999ae22c · outbound

This paper cites A correspondence between Bayesian estimation on stochastic processes and smoothing by splines,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning A correspondence between Bayesian estimation on stochastic processes and smoothing by splines,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.377606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.377606Z digest=sha256:ee333eb887bd0ca2ebba47367523454edac7e78adfc857f1078f1ce264d5ccb4

Observation 1044a165-ffc6-4346-9b63-74a4326bb042 · outbound

This paper cites Universal approximation capability of EBF neural networks with arbitrary activation functions,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Universal approximation capability of EBF neural networks with arbitrary activation functions,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.493294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.493294Z digest=sha256:dfbef10d4044a68376f38bdbd28cf948ca5d07e5c14790c9a8265ae2f946a045

Observation bfb26320-5c77-4a12-b487-70a56b7d1fb9 · outbound

This paper cites an unresolved cited work.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.597574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.597574Z digest=sha256:c5eb8fabaa3dd5f0c7ee213db442fdbb2ebdb2fff5cc71956b4aba789f04d2c7

Observation 32153aa7-e41c-4793-a8ea-64a7c885a306 · outbound

This paper cites Williams,Probability with Martingales.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Williams,Probability with Martingales

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.778407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.778407Z digest=sha256:0141a30f82836bad32a93cc30ef7ab9194fdb81b1daa380585a588efcc6f7424

Observation bcf3dbe5-4f35-4d4c-85f1-42e801afae75 · outbound

This paper cites an unresolved cited work.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:41.921653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:41.921653Z digest=sha256:36b6c1fdcbb283dd6dc7af90ef1cfc713b41372d273bc03ca9b8f4c94a1ab356

Observation 190f0118-1b9d-4d9a-bf30-c465cafb6e52 · outbound

This paper cites Rudin,Real and Complex Analysis, 3rd ed.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Rudin,Real and Complex Analysis, 3rd ed

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.036054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.036054Z digest=sha256:00f446b6e8a101b8c4399f596eb3f70c03bc86781e7ed8697fc08b59d401055e

Observation a42f425f-f0e7-4518-807f-819e714a0b7a · outbound

This paper cites Wasserstein Riemannian geometry of Gaussian densities,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Wasserstein Riemannian geometry of Gaussian densities,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.261238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.261238Z digest=sha256:3ffa9ed31f529771f23d7765b73329929ba629d4145d6e70e662668c533b0f25

Observation d2a32366-ddf7-4e9a-a8fa-cb5ca0bb7bf5 · outbound

This paper cites A sampling approach to finding Lyapunov functions for nonlinear discrete-time systems,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning A sampling approach to finding Lyapunov functions for nonlinear discrete-time systems,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.382131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.382131Z digest=sha256:3ffb0c75b241afe5ad13d6646f00afc670557c5e6bcc8c4fa4adb451e4bedf1c

Observation ccb3f4e8-2daf-48ef-8b40-fb6155c7b56b · outbound

This paper cites Boumal,An Introduction to Optimization on Smooth Manifolds.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Boumal,An Introduction to Optimization on Smooth Manifolds

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.546000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.546000Z digest=sha256:1b24da579aafa57dad355219f89ce4133b529086e603eb60a60ab154b98074a1

Observation 9b0c3880-bc09-43f4-a4fe-b2edee3979d9 · outbound

This paper cites Early stopping—but when?.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Early stopping—but when?

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.672251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.672251Z digest=sha256:e6aec5d6ec8e508e5fe447f0bcb1b6fa295d4d9e281fcf77905f1d0956df8524

Observation d10912af-6da8-4f2e-97ba-1343143a56f5 · outbound

This paper cites Learning near-optimal poli- cies with bellman-residual minimization based fitted policy iteration and a single sample path,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Learning near-optimal poli- cies with bellman-residual minimization based fitted policy iteration and a single sample path,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.814788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.814788Z digest=sha256:427523d22e87686519ce9941a87fdd34c0b123b3feaa0f0f969c760e8eb332e8

Observation 65a8e296-ff23-403a-942e-ee2cbd1726d3 · outbound

This paper cites Deterministic Bellman residual minimization,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Deterministic Bellman residual minimization,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:42.924151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:42.924151Z digest=sha256:fd75794dc634ee49f53e141321bab3d2feed0b7f2b09293fa5bb9eee88ec0fa6

Observation 4f0b9e33-f2f4-4e64-ad0c-6e5d7e05978e · outbound

This paper cites Big Omicron and big Omega and big Theta,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Big Omicron and big Omega and big Theta,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.034296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.034296Z digest=sha256:124be68fc786eb8c36cdab753f330f28926c0ab140ffa724c5e0293af8dd10a2

Observation 259417fb-cbfb-4951-b447-39aa192f71d2 · outbound

This paper cites Reinforcement learning in continuous time and space,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Reinforcement learning in continuous time and space,

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-08-03T15:00:43.137851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.137851Z digest=sha256:d361d16f077049bb9949f222b2fcfec7fb5cab0d393bbb12816867d77cfbb626

Observation d99ac932-6bb8-435e-8319-4bcf81e86838 · outbound

This paper cites Efficient memory-based learning for robot control,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Efficient memory-based learning for robot control,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.266206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.266206Z digest=sha256:deb7cc103da493a7540e93bfbeb8b1def512775b9c8604980bc6e2d00bd71873

Observation 4a065fd9-f530-4c43-a003-d986e25616de · outbound

This paper cites The swing up control problem for the acrobot,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning The swing up control problem for the acrobot,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.428927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.428927Z digest=sha256:926b62b2e0c126c04d07eeee2aed1a7658198119afcc573a9eb5ceff4ed699b4

Observation dbe09ffb-b129-46ca-85df-6ce287e156c6 · outbound

This paper cites Dueling network architectures for deep reinforcement learning,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Dueling network architectures for deep reinforcement learning,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.546416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.546416Z digest=sha256:cff1bb6c993370862ec8a759e7257ec060f5c8e7c398f98b5d0a372d4d9e5d74

Observation eb698973-f46e-4815-8eed-74d5a33ce3cd · outbound

This paper cites Proximal Policy Optimization Algorithms.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.650378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.650378Z digest=sha256:b85f47403f2df574da8e8ccefff52c6a26427908da7d42c19175c20dc1394584

Observation fd934d2d-3976-49cb-8954-c613fe2fbaac · outbound

This paper cites Julia: A fresh approach to numerical computing,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Julia: A fresh approach to numerical computing,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.812659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.812659Z digest=sha256:6b7fb3820b6301e4d7d1030c8d3377c7095c67bacd779d2b30905dc1a18f0a29

Observation d31ccf1c-9e64-4f23-b859-fa24d7b7b9c6 · outbound

This paper cites Robust reinforcement learning using least squares policy iteration with provable performance guarantees,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Robust reinforcement learning using least squares policy iteration with provable performance guarantees,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.900090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.900090Z digest=sha256:6e092683edb2ed4f332ae1c341cc2dbb7f70e995e917bcdb224a9248375f77c9

Observation a3fc58a1-2470-4994-a10b-ca3c41af1a21 · outbound

This paper cites Geron,Tiny-dqn, https://github.com/ageron/tiny-dqn, 2017.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Geron,Tiny-dqn, https://github.com/ageron/tiny-dqn, 2017

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:43.968185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:43.968185Z digest=sha256:ea9d74268f9f4eaa8f0c211b1707757d0c68ef77f8871db2469efefa4e7bbb09

Observation b819367f-3f15-46ff-ae46-43a413d06ed2 · outbound

This paper cites an unresolved cited work.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:44.063603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:44.063603Z digest=sha256:dd2585528c1a9ffab86c15483bb8c4da3e80a7029d77a0216da13cd94ec07dca

Observation 59f12b05-1702-4d05-8162-2e080e7b0234 · outbound

This paper cites Density in approximation theory,.

Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning Density in approximation theory,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T15:00:44.145471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:00:44.145471Z digest=sha256:90ba2279c4b396e8d5411fd69f32460a2bb62dd75440083207465abe3b816da0

Pith citing papers

Observation ef344002-af4d-4b95-a275-c1b983409b7e · inbound

Sparse Gaussian-Mixture-Model Q-Functions via Hadamard Overparametrization for Online Reinforcement Learning cites this paper.

Sparse Gaussian-Mixture-Model Q-Functions via Hadamard Overparametrization for Online Reinforcement Learning Gaussian-Mixture-Model Q-Functions for Policy Iteration in Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-30T21:23:27.894346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:23:27.894346Z digest=sha256:f7d27c2f1a122e6cc90cf624be5bd23eda80211e3914043543d13a4479d289bf