Pith. sign in

Paper Citation Record · LEDGER

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning

As of 11 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2501.10605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10605 v2

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:08:00.458822Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T02:07:54.718436Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:26:56.835415Z

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 27e86555-2df3-473c-8b85-1e108c2caf0e · outbound

This paper cites Reinforcement Learning with Wasserstein Distance Regularisation, with Applications to Multipolicy Learning.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Reinforcement Learning with Wasserstein Distance Regularisation, with Applications to Multipolicy Learning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T19:08:00.726131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.335175Z digest=sha256:87272c221a402a2df6559111bdac7de666b048ea3d6b09339c7ed675042451c6

Observation c41fbd83-e1e5-4e25-9a5d-aa8cfeb68377 · outbound

This paper cites Wasserstein Robust Reinforcement Learning.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Wasserstein Robust Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T19:08:00.342772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:08:00.342772Z digest=sha256:2e20c4ed81d789beb9062cc7b366f9d7a91d7a6954f15fee474f99e79771de94

Observation 80209321-f39f-48aa-8f67-f003149ed159 · outbound

This paper cites Risk-Aware Reinforcement Learning through Optimal Transport Theory.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Risk-Aware Reinforcement Learning through Optimal Transport Theory

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T19:08:00.348723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:08:00.348723Z digest=sha256:ea4c5040b5487760e746d84e9a85615cd2497dba8aa631ef6183c4cb60ced411

Observation 783ceb5d-8bb9-497c-9cf2-7af5edd25edc · outbound

This paper cites The Synergy Between Optimal Transport Theory and Multi-Agent Reinforcement Learning.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning The Synergy Between Optimal Transport Theory and Multi-Agent Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T19:08:00.354941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:08:00.354941Z digest=sha256:9ba5d260ffb239ffa1499f60e6446480fc44c89d70e9fd59f4ca2985384c298b

Observation 390257c5-a4ee-47ef-a04d-169412e44573 · outbound

This paper cites Finite-Time Analysis of Entropy-Regularized Neural Natural Actor-Critic Algorithm.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Finite-Time Analysis of Entropy-Regularized Neural Natural Actor-Critic Algorithm

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T19:08:00.624139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.360475Z digest=sha256:5978b8a7c1b2a5ffcdf602bab9349428b3e4c8f684f2ffda567bd5c095b961bc

Observation cca2d674-2f30-4142-80fa-50949992e138 · outbound

This paper cites In: International conference on machine learning.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning In: International conference on machine learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.935489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.366350Z digest=sha256:8faed9485f90e2bb262561a0ace82e6ab6c44febc4b1f83bf648711302f8e223

Observation f567663e-fb89-4b3a-8d3b-ff9a89e36f4a · outbound

This paper cites IEEE Transactions on Systems, Man, and Cybernetics, part C (applications and reviews)42(6), 1291– 1307 (2012).

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning IEEE Transactions on Systems, Man, and Cybernetics, part C (applications and reviews)42(6), 1291– 1307 (2012)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.918792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.379061Z digest=sha256:2d91b5cef01109a147c95914694af4f0e4c17bcc73acd355705806d11f758aab

Observation 9c7b3e0c-3a48-42bd-b1fd-e1ff2134626b · outbound

This paper cites Mathematical Finance 33(3), 437–503 (2023).

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Mathematical Finance 33(3), 437–503 (2023)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.898795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.385215Z digest=sha256:ea012fc82573673c602a10c7e6eeebd724375c862aa8c0fe53ebb819473d6069

Observation f04ea8b2-9071-4107-99f7-74ce89813bcf · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.875726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.391373Z digest=sha256:8131caa318948b149f305364fad6608570c91b5da3d48fc75158b1ef9b59ce67

Observation 011c81d1-7518-4cb0-97ae-24d542645976 · outbound

This paper cites Robust Reinforcement Learning with Wasserstein Constraint.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Robust Reinforcement Learning with Wasserstein Constraint

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T19:08:00.396692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:08:00.396692Z digest=sha256:581959b29cdc0fd9e07a2e0471a5287b74901bd5e69d7924f83837e7fefbc89e

Observation a1375f69-eae0-44be-85b7-2d2427b22b73 · outbound

This paper cites The International Journal of Robotics Research32(11), 1238–1274 (2013) Wasserstein Adaptive Value Estimation 13.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning The International Journal of Robotics Research32(11), 1238–1274 (2013) Wasserstein Adaptive Value Estimation 13

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.857858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.404321Z digest=sha256:9548e8549dce3aa2dc6204d9d6975d4fa5297f7152335619b084b58043018480

Observation 0c480420-d8c3-4aa7-840e-819308682cad · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.835770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.410129Z digest=sha256:e017927258c87bf029298ee59d133762b158bdb8de37d723db4d3d560c9cc8c3

Observation 6fa6680e-cd49-4565-a6e5-e58e7efb37c5 · outbound

This paper cites Advances in Neural Information Processing Systems 32 (2019).

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Advances in Neural Information Processing Systems 32 (2019)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.816073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.416693Z digest=sha256:9888eeb7462a57b9d7a97488cd3188482825ffbe7d3f76240d2dbb162472dad2

Observation c85bfb0c-69c5-4032-b2cc-9d5d7591f883 · outbound

This paper cites Engineering Applications of Artificial Intelligence136, 108911 (2024).

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Engineering Applications of Artificial Intelligence136, 108911 (2024)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.796877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.424703Z digest=sha256:175c9628a58f5863469cec68c46160cd15a5f1a58374c61b4b8dd7a0b4101247

Observation 82b555b4-9546-46d4-9335-ed480bbb0e6a · outbound

This paper cites On Wasserstein Reinforcement Learning and the Fokker-Planck equation.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning On Wasserstein Reinforcement Learning and the Fokker-Planck equation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T19:08:00.431086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:08:00.431086Z digest=sha256:27309b104d20f4edf5a8fe31f3ee933bc6ca0d01441136e6dbec647e59321229

Observation 3179028c-dd73-4255-b33a-6a3a00d6ab63 · outbound

This paper cites Optimal Transport-Assisted Risk-Sensitive Q-Learning.

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Optimal Transport-Assisted Risk-Sensitive Q-Learning

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T19:08:00.541078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.442135Z digest=sha256:5b8b377c33c460e8cdb60a85dee8342357754878cd2737079e4ce148c10addc0

Observation 2860e3c0-8847-497c-95af-8bee2fa9848d · outbound

This paper cites ACM Computing Surveys (CSUR)55(1), 1–36 (2021).

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning ACM Computing Surveys (CSUR)55(1), 1–36 (2021)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.776438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.451662Z digest=sha256:2512929808940d51652a1a1850ba25eeb3f506ce167bbd9a0bf0e54ab1ac7f09

Observation 1e5e5239-b5d4-4090-9335-b834bf49f898 · outbound

This paper cites Advances in Neural Information Processing Systems34, 15993–16006 (2021).

Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning Advances in Neural Information Processing Systems34, 15993–16006 (2021)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:08:00.756321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T19:08:00.458822Z digest=sha256:f61000c3a31aa10dc0fcc487c13fc86537e51c2423fac74d693d616bacc36d27

Pith citing papers

Observation 74577a06-21c3-4cb3-89e7-1209ad3cad60 · inbound

Geometry-Aware Dataset Condensation for Diffusion Model Training cites this paper.

Geometry-Aware Dataset Condensation for Diffusion Model Training Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:56.837015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T02:07:54.718436Z digest=sha256:a566652fdf32eb84b1267471f7a3998256889336f83ec4e477dbdc11ef51018c