Pith. sign in

Paper Citation Record · LEDGER

Policy Newton Algorithm in Reproducing Kernel Hilbert Space

As of 8 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2506.01597.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01597 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:47:50.149885Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b7b9f4c-72a4-46d2-b80e-6957c595ad6f · outbound

This paper cites Practical kernel-based reinforcement learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Practical kernel-based reinforcement learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.956885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.631890Z digest=sha256:0fa6ef48630a30a8395ee468f26f0970fb38467b615a472042680e3f7a7e8bc8

Observation 3653c063-a122-4e06-be01-65c5a2d96b93 · outbound

This paper cites Optimization methods for large-scale machine learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Optimization methods for large-scale machine learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:59.667734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:59.667734Z digest=sha256:e1c8bc61c3f184cfa2959adf16a579c3cc944f68c0f53318cbe2761126413be1

Observation 38d48e19-5543-48d1-9557-ebf4977680f1 · outbound

This paper cites A tutorial on support vector machines for pattern recognition.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space A tutorial on support vector machines for pattern recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.876836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.701880Z digest=sha256:d4ca3f6026ead2632444de6efa69498ec0014f056c592e2247aea823cb9b513f

Observation 13bed701-f904-4d3c-a4ed-4720ea77e7aa · outbound

This paper cites Second-order kernel online convex optimization with adaptive sketching.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Second-order kernel online convex optimization with adaptive sketching

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.847498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.745345Z digest=sha256:410ab9fa08c349a2eb8888dfb723428e0a4ef056285e6d7785d06f3f5503b5d8

Observation 833dadd6-932b-483e-a109-ac35a1e76fa7 · outbound

This paper cites Efficient second-order online kernel learning with adaptive embedding.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Efficient second-order online kernel learning with adaptive embedding

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.817517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.790905Z digest=sha256:21f6a71fb5e40b701d0821de4617e2e12ae02e9ee585cdc4aae896fc734680ed

Observation 66a6884c-38b4-44ff-b2d4-ff98d238c5bd · outbound

This paper cites Super-universal regularized newton method.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Super-universal regularized newton method

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.787831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.844612Z digest=sha256:453e84ca943089c8ebc12a17407036a3f1b3fcbee280b82399aa0ee85a538f38

Observation 3d5ebd50-3cb1-4545-8c89-e026517115d8 · outbound

This paper cites Approximate newton methods for policy search in markov decision processes.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Approximate newton methods for policy search in markov decision processes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.756638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.880899Z digest=sha256:171cceeb467de7e52797b4463724ebdb41f883930801978d39d9a4afe65e1be0

Observation 97fcdeef-69c1-4c48-8455-3e7928ef8eb5 · outbound

This paper cites Stochastic first-and zeroth-order methods for nonconvex stochastic programming.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Stochastic first-and zeroth-order methods for nonconvex stochastic programming

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:59.916768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:59.916768Z digest=sha256:5149ef88b028a9be65cf167836aa61db7038b401d80b04b69cb4f826781124ec

Observation dbd4ba99-2f8b-4d8e-9153-fce6679a0e5c · outbound

This paper cites Quasi-newton trust region policy optimization.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Quasi-newton trust region policy optimization

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.708524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:46:59.963385Z digest=sha256:0fb71b3621c223e919773affa5dc209c67f4c591f08f81569fa93dc28458caf5

Observation f4079ccf-9e7c-470a-acf0-0f3c0e83f18f · outbound

This paper cites Convergence and decomposition for tensor products of hilbert space operators.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Convergence and decomposition for tensor products of hilbert space operators

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.677002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.000150Z digest=sha256:f8d2f76968f4a77347eb789d980fabc03e1c9a57cefb06ff0615230f10e71ff3

Observation bc939fe8-1343-4f21-9533-c852ae7b9509 · outbound

This paper cites The conjugate gradient method for optimal control problems.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space The conjugate gradient method for optimal control problems

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.637324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.032444Z digest=sha256:0c0836884d075ab90e1831c87cae1361fd1e196b1e9de0d2fccf25e12bc40fd0

Observation fe4ca90f-4ebe-4a9f-9738-6534b912330a · outbound

This paper cites Fastfood-approximating kernel expansions in loglinear time.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Fastfood-approximating kernel expansions in loglinear time

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.608166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.067926Z digest=sha256:8198a19b41c78a7ff9327c1881f5188d5d02ffc1ac740558547a51d51c286181

Observation 4b84e4da-8778-4e93-95e0-27ebad03447e · outbound

This paper cites Dynamic asset allocation exploiting predictors in reinforcement learning framework.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Dynamic asset allocation exploiting predictors in reinforcement learning framework

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.572162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.111348Z digest=sha256:5439883e3e2b66db26b3611a4e2f58d32e610a8550849ab2dd96daf476b65f21

Observation e1d471c3-68e0-4839-859e-f1e1a4992d63 · outbound

This paper cites Parameterizing non-parametric meta- reinforcement learning tasks via subtask decomposition.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Parameterizing non-parametric meta- reinforcement learning tasks via subtask decomposition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.529875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.149331Z digest=sha256:cc245a0f2b50627011d8f3752b63b5bef2c810f18792fcf23fa428ebdb2d7484

Observation f8f600a6-3d46-4ec1-8724-ed06dc5ad401 · outbound

This paper cites Modelling policies in mdps in reproducing kernel hilbert space.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Modelling policies in mdps in reproducing kernel hilbert space

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.488772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.183973Z digest=sha256:9fdcc713ce5add119cdd35c5426dc91eb296e7623b267b92f0c094d9ef830fed

Observation 6b61afb5-dbb2-4525-9a9a-987a1dcc4bd5 · outbound

This paper cites Approximate newton policy gradient algorithms.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Approximate newton policy gradient algorithms

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.454621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.220120Z digest=sha256:0732349ccc25122c881e218de3c838b79f644bc2341e63a9ad5264cc53e5c99e

Observation 7f25bcc2-dac5-46eb-acdd-6250d14cb206 · outbound

This paper cites Large scale online kernel learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Large scale online kernel learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.421894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.264079Z digest=sha256:238400c6d9e866855d4a04c8ea1987a7c6a2ea3b29a070ebd8643eedb6e59383

Observation fc1380d5-8eb6-40cf-8ff1-2058ef874c43 · outbound

This paper cites A cubic-regularized policy newton algorithm for reinforcement learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space A cubic-regularized policy newton algorithm for reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.391581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.307378Z digest=sha256:6f10cfd2e147b0c338d3fdb54d9c367ea9bf0733f4eeccd067909066449f735d

Observation 7231c35e-353e-4fc9-96fe-a75b203814e7 · outbound

This paper cites Methods for calculating fréchet derivatives and sensi- tivities for the non-linear inverse problem: A comparative study 1.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Methods for calculating fréchet derivatives and sensi- tivities for the non-linear inverse problem: A comparative study 1

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.356665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.341786Z digest=sha256:4805972a9ab2db7bd691e48c4e4bd00a7c2923acdd3c8bf67fcce737c17f1487

Observation 5ea5a9df-5caa-42eb-a7bb-bb25694c7b8b · outbound

This paper cites A scalable kernel approach to reinforcement learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space A scalable kernel approach to reinforcement learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.323716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.378817Z digest=sha256:dfb5ed270bf489af7262d41eda48e9b015b3c022e0397bb420a42b315e71cdb6

Observation 8bb6f7b7-f84c-4c9d-a991-44f6b2920cc5 · outbound

This paper cites Nonparametric return distribution approximation for reinforcement learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Nonparametric return distribution approximation for reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.292540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.422894Z digest=sha256:dc0096c2621d4c44818204cb562484f1e7e4499bd2fb0d7e7dcb459d31d574d1

Observation 3329dfd0-b5d1-4e86-ac10-4ebe278ef292 · outbound

This paper cites Universal approximation using radial-basis-function networks.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Universal approximation using radial-basis-function networks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:48:04.257731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.458400Z digest=sha256:e4167366ae6fd911cc404629a9e88ce3cbcaa41bc16e6d6dcdd54c1adf58ffc3

Observation 003b426f-f54e-4705-85af-70a930919bdf · outbound

This paper cites Stochastic policy gradient ascent in Reproducing Kernel Hilbert Spaces.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Stochastic policy gradient ascent in Reproducing Kernel Hilbert Spaces

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:52.213074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.509650Z digest=sha256:aa0890f7d355f62d6bb56a91c24ec709f32c066e03c9e49a58adace0366ac13f

Observation 81c94938-e999-4a9a-9c0f-27865210e575 · outbound

This paper cites Policy gradient for continu- ing tasks in discounted markov decision processes.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Policy gradient for continu- ing tasks in discounted markov decision processes

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:52.122310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.567753Z digest=sha256:25f260fed54b0e8be14568705b2b7dd68dc908d321c629fd7c7a11f3acf76ab4

Observation a91b2baf-e769-48e9-8fd6-0c563ae65891 · outbound

This paper cites A generalized representer theorem.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space A generalized representer theorem

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:52.016067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.600731Z digest=sha256:f2127eb8a8de03cc42d6defaeb4a252338d5d75ed09458351251fe97f74f6c35

Observation 23626f5e-bdab-411a-8479-c0da69601e11 · outbound

This paper cites Hessian aided policy gradient.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Hessian aided policy gradient

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.935781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.637328Z digest=sha256:7b7922413cf9ff027e1b9a00b96a1e52707f7bbb61e4eea35677a32eba8cd1f2

Observation 721a06b7-58e8-412d-9f64-e2b8cc9f8b8a · outbound

This paper cites Sutton and Andrew G.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Sutton and Andrew G

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.879277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.668055Z digest=sha256:b6131a909272f969bf92c633fbe014700a82527534f4c1a689f0e01265e447eb

Observation 780da708-98aa-4f7f-9af0-d8ba102221df · outbound

This paper cites Policy gradient meth- ods for reinforcement learning with function approximation.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Policy gradient meth- ods for reinforcement learning with function approximation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.811639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.703265Z digest=sha256:6ac97ff44bebe8c80207f8865852a47e290216cef90b69b619d1ff25631fbd7a

Observation db8fda62-43a1-451d-be2c-f9c699a1f68f · outbound

This paper cites Mujoco: A physics engine for model-based control.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Mujoco: A physics engine for model-based control

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.659982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:00.747483Z digest=sha256:b124a2cb64d585457cd30fb7058f45f8d8a7b8c442410e0038dd8388427f1822

Observation a5b2a4c8-3792-4df8-aa02-fb2cbf635817 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:47:00.789638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:47:00.789638Z digest=sha256:304ea0df17bc3cf453b0bf424ec5a324b7b91bb5870bd1e9ed8904c7106c9449

Observation 02af97c1-5ca6-4bb0-8782-5995bc1b82e6 · outbound

This paper cites On the convergence rates of policy gradient methods.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space On the convergence rates of policy gradient methods

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.488017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.033620Z digest=sha256:d09147aed3306125b78cac218f01ca77a48d778562ef289b03f145cf15cfc056

Observation 19576e37-37c5-473e-b4d1-e447101462d8 · outbound

This paper cites Safety aarl: Weight adjustment for reinforcement- learning-based safety dynamic asset allocation strategies.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Safety aarl: Weight adjustment for reinforcement- learning-based safety dynamic asset allocation strategies

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.261797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.594886Z digest=sha256:bfc862fb65b6fbdff89482031eeb2176144858c16a894dad2f1d4c41a2ed8998

Observation c966317b-dbda-4ec6-aa2d-c657b80d4d5e · outbound

This paper cites Global convergence of policy gradient methods to (almost) locally optimal policies.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Global convergence of policy gradient methods to (almost) locally optimal policies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:51.088717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.666149Z digest=sha256:bed169aa4a8925e1776db66968ef55b1107e851140f7231427c7c327b4c4b100

Observation f6117b3b-e7c0-458c-b880-c7702b497bf9 · outbound

This paper cites Residual kernel policy network: Enhancing stability and robustness in rkhs-based reinforcement learning.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Residual kernel policy network: Enhancing stability and robustness in rkhs-based reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:50.938637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.723203Z digest=sha256:3b99a4bb93782efc02f805beaafaac581ddc590daf5aae9e82c9396d5c930c81

Observation eefec5d7-cb49-43c7-8f8c-d4fe312d2f92 · outbound

This paper cites Let Xl(¯hα) = ⟨∇h log πh(xl), ¯hα⟩.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Let Xl(¯hα) = ⟨∇h log πh(xl), ¯hα⟩

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:50.866789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.780070Z digest=sha256:c45440a312617310b440868dbb5922dd70ad440a1c91905a5a7c8b49727d548d

Observation c17661dd-2369-47db-8a2b-8af870177b72 · outbound

This paper cites The quadratic form is ⟨H (2) op ◦ ¯hα, ¯hα⟩ = PM l=1 Ψl(τ )T Cova′∼π(·|sl)[K((sl, a′), ·)] ◦ ¯hα, ¯hα.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space The quadratic form is ⟨H (2) op ◦ ¯hα, ¯hα⟩ = PM l=1 Ψl(τ )T Cova′∼π(·|sl)[K((sl, a′), ·)] ◦ ¯hα, ¯hα

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:50.761713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.833573Z digest=sha256:ff0bba8e2d96e4b2072eb0d2feb0948cef84821882638be0bb5fb2afa6d0f7bf

Observation 79e304ea-8a78-47ab-bb4b-31c99d707f03 · outbound

This paper cites an unresolved cited work.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:47:50.668667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.917454Z digest=sha256:5a4f13dc26e3f33e89af903123365c589f7e121bdcf2f235ae0e33c2ea917df3

Observation 6c83b39f-e70e-4ee4-9eed-320ffc79e778 · outbound

This paper cites an unresolved cited work.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:47:50.586986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:49.988158Z digest=sha256:e00dc3a6b26aa8a68f37b203a71395e263c5deb81fd083b3dd94ac038a84e83d

Observation 01608c29-4aaf-4239-bc34-65ba0de52e6a · outbound

This paper cites The conjugate gradient method is employed to efficiently solve this linear system, avoiding the high computational cost of directly computing (H + β 2 ∥ ¯α∥I)−1.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space The conjugate gradient method is employed to efficiently solve this linear system, avoiding the high computational cost of directly computing (H + β 2 ∥ ¯α∥I)−1

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:50.492623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:50.046535Z digest=sha256:407ad3cf6d9b695d18432b10d6abf02faec1ca77fe2b54c2377d1548cc250354

Observation f067549e-a9fe-47fe-ae8d-823daa821abe · outbound

This paper cites an unresolved cited work.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:47:50.408207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:50.080720Z digest=sha256:0c0157cd25738992e0520bc93a663b0f62f43c1ee0dd3e9af45235b3c75931a7

Observation 0da68224-54a2-447d-80f2-e8747929df11 · outbound

This paper cites To balance optimization accuracy and computational efficiency, we set the convergence tolerance to10−3 and the maximum number of iterations to 500.

Policy Newton Algorithm in Reproducing Kernel Hilbert Space To balance optimization accuracy and computational efficiency, we set the convergence tolerance to10−3 and the maximum number of iterations to 500

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:47:50.310407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:47:50.149885Z digest=sha256:b3555b5452555d62adfa52a1e4247b1339592290aab8faae577d40367a3682a7

Pith citing papers

No inbound Pith citation observations are available.