Pith. sign in

Paper Citation Record · LEDGER

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control

As of 23 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2605.26361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.26361 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T21:10:16.593076Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved38
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2444ae96-03c6-4120-b7af-a37ee03a5a7c · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:301182f21115c97d4da2d18d83bbed4bf107261cc134c8c7bb27a2f7de3408f9

Observation 2c7b13ac-4fdd-4818-9692-f1c1ad966a66 · outbound

This paper cites and Wager, S.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Wager, S

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:1220bcaaeec8c44ac682806dd131930c8f985c0094abbb1b47ca949f439d5a17

Observation 900dd0ff-5a3f-459a-8321-e626ee63c8d4 · outbound

This paper cites and Tsybakov, A.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Tsybakov, A

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:4677d201f2336a23e74fbb25b55a4c5ad86774ac9145babb964dc5e97fab97bf

Observation 632335b0-24e1-479b-bb59-951f48bf90d2 · outbound

This paper cites and Rudin, C.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Rudin, C

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:a7c5cbe8ae7c6f5a28cf4170927c320e91552281079bae80fbd141e3f9f46638

Observation 4e3ffe05-6c0f-4626-8c52-48097bbbd98b · outbound

This paper cites L., Jordan, M.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control L., Jordan, M

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:68f124a92fef1a1b5d28a9c7249429bb22b10c6b6c05444d8244ad2b768f8e09

Observation 9a053025-f8e5-44cd-a669-8d5c3ac99859 · outbound

This paper cites (2012).Dynamic programming and optimal control: Volume I, volume 4.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control (2012).Dynamic programming and optimal control: Volume I, volume 4

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:3f29eb5dd62560d44775aecda8874d957d3dc64b56f3c6132f46d5e64d2362a1

Observation 2936a192-7239-4f61-a90b-321ec352c887 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:97db4ce53e5316e1426bf167354bdfb4fb65546f4c3633c489d2272809edf774

Observation 57ecbe1d-88e6-49e5-bdf2-d5fa0b739351 · outbound

This paper cites and Kallus, N.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Kallus, N

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:71671ca635eb69c978d3f310249ad61d6e54dd7b38ea86369fcaa021cfea959e

Observation 90c3cb8e-72ad-46bd-ac5b-38181459c632 · outbound

This paper cites and Zeevi, A.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Zeevi, A

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:e1ff2de550b3dd9a8af9d93d92ead4345567421960fda831b5beb9472bdca749

Observation 92c9af2e-6ce7-4fea-a7b8-451067f785a2 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:22c756077abb48ebf2407ce06d2c5e23272b2b5867d2d706fdb727784cad9399

Observation 7f1504ba-3634-4b13-8d20-1a48f3d09ec7 · outbound

This paper cites predict, then optimize.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control predict, then optimize

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:fc59465fe56bc4b5a901afbaa48cdc2ea717da74ba3d962372210c1697ac773b

Observation b2377e80-f750-482d-9ba3-cf12a2d9eed9 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:fd96f3a1397081bb40194d28bb01d0b2145e5cc2f3249c08d293737ea1b7d983

Observation 1192fba8-8f97-4a95-84ab-5765c44419be · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:430c6aabaa02d451227dda8453a0e2d43ca01f6a9ec3ecbc2324ec17f46e0b98

Observation 84ecbff8-46e9-44c2-9946-16d1d9532c94 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:e029644cc893a656c1c644737c2719ff76d93c45f31fbc410ac3f5ef5db4fb65

Observation 2cb8b9b1-9ea9-4748-b629-1bb428fd500b · outbound

This paper cites and Nickl, R.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Nickl, R

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:d7bc31cd74a988fe26513301ea8d3b027955aad203c46dfcd651dba4a49d96bc

Observation f4f1f95c-dfc1-47c7-a9d8-4c0950fd6437 · outbound

This paper cites (2012).Adaptive Markov control processes.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control (2012).Adaptive Markov control processes

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:7459f11aa620b7250814a9d885a29b74c2bb3e2a3bff63f8aa3abea3fe17e48d

Observation 4350c4b8-819b-4c29-a6b8-5825a51d5d92 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:01c489cb4178ffafbd18b67afcf453405e04d0f0624e5f1b339aae8b9fe777d2

Observation 2e875c27-c024-47ae-b0d1-fbea3c42b89d · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:eef16c487ed4754ab455185e0356afbeaca1cadb615371531424624811b5a5e4

Observation 7c5e5ae9-3b01-43a6-8857-fef41a6cc590 · outbound

This paper cites and Mao, X.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Mao, X

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:c86e636f5e3072c23a83b5c13595d549e84b91a7772c526a770f2e9f9dac8770

Observation 48b458a4-62a5-421e-95a9-a96cef901524 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:2254ab2d10a68fbda58ad3f06d3cfac03df51d89d58dac0261c444d008a07158

Observation 1cd0ee8f-4b5a-4fe8-b948-97c03b6632aa · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:cef6f41579da632ca9853ee6d4c5a95971a8bbc807e21c8e4a1b4e3682e18f9b

Observation d9bdfe93-6406-4af9-8287-d9c7edf136f5 · outbound

This paper cites J., Shapiro, A., and Homem-de Mello, T.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control J., Shapiro, A., and Homem-de Mello, T

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:b3a5e6952e0f5547866a6db66f94d7d661235226521bdb064e38d843f7a48964

Observation e75dd49f-ed32-42c9-af18-544d14567c22 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:923302c9766a9b6455156739e1af98645fc076991f760e2f36d2c789a8840bb3

Observation 7629d5e1-ade2-4d45-864a-d7e78e5fad3d · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:14:00.563572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:3b243c0e48372a3f1e333a0c6ee821989bcb424416df2bbd306e5fa66f256515

Observation a7e1af88-6671-43e5-9190-a999463f1c0c · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:3aa7044cd2a206e70cd3fd6847cdbc7f713b32ff8c69ee419a61b3e8a78ef8fd

Observation 0487fb79-d34e-4c67-ba40-256fef8e5259 · outbound

This paper cites and Chambaz, A.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Chambaz, A

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:8fd340ee79a3a54d29082a5b771a7491b3b0a0d0aa75f31da01719948b54910c

Observation 87be002a-5119-40ac-9f63-92e93c376893 · outbound

This paper cites and Tsybakov, A.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Tsybakov, A

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:3b36f600381c96105496b6435093a5d2d08080129cb0e46ceabfffc2ec8ac872

Observation 992f227e-73a7-4466-be2b-27bbe25c66b1 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:b5014ab121f7284c54850184452ddb977cd3636dda541fc735c5b7cdf54b66a6

Observation 6b763517-9847-455c-be2f-f8ebfa5a06a0 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:842d928289547c1b9533cd8ab4b1ec787b20793acbb9e3e74b28d8863c1bf1e0

Observation 4eb129ca-b264-4e16-b434-4f5358345690 · outbound

This paper cites A., Veness, J., Bellemare, M.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control A., Veness, J., Bellemare, M

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:add3865ccce26c6e965baf70af4bafbf41de061d3d6ff861569cbe4b0fc1dd64

Observation 987244b0-d653-49e9-b70d-eb01248e995c · outbound

This paper cites and Szepesv´ ari, C.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control and Szepesv´ ari, C

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:6a6b43304e0c756ab92894e62eb54612731e8157293bceceaead9920d2d6d53e

Observation 4385d703-1a70-4a56-866a-0197d4b7e231 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:872005c7e7c26f0ff1d1fa0424e9bee217989110b367603b06ef3748be5e2451

Observation 63a35e8a-416f-4b80-84ef-6b86f4127db2 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:a45c34537913b2ac8e005deb1b59095e9f2b392ac9669272ec898be5febcab87

Observation 29034b05-e63a-4625-8cbf-49e3de76f5a8 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 34

Resolution
parse uncertain
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:1b58695cce731ddd5e456fafc3ad2efc60aa64cf0943ef687f433670a0ec7f5e

Observation 48ae62ab-7dc8-417a-b303-4b92465647b4 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:edd31d975e56e93e840daf04a6d7c3b2ce36e3053863bafe8a756ae09620d491

Observation c51380b9-6dc3-4fff-af29-54ae4102ef81 · outbound

This paper cites (2021).Lectures on stochastic programming: modeling and theory.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control (2021).Lectures on stochastic programming: modeling and theory

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:7d65273c398ddb6a8a53f82dfc6249a39bbc4b186d80d82fc08d94cfe50aa464

Observation 9d3cf721-60b9-4230-a08a-f580a7694d15 · outbound

This paper cites S., McAllester, D., Singh, S., and Mansour, Y.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control S., McAllester, D., Singh, S., and Mansour, Y

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:a7d07533b5c9a475b0ef1a6628c5082be5546c3130bdcf562eb2355c89353217

Observation d6392a54-bb2d-4a64-a38c-896d69d60289 · outbound

This paper cites (2008).Introduction to Nonparametric Estimation.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control (2008).Introduction to Nonparametric Estimation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:e7fb690ce3726b6e75b926288db0fe2abed901f148636eca8c2120eeef174dea

Observation c13ebe4a-2605-4618-9cc7-149680ccce6b · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:98f73b8905db3430fe9e00c36e4812e51b973b7a613fca6b22388169230f981b

Observation 8488c31b-851b-4e08-8701-12b442c0b87b · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:14:00.572083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:5911962b1c0d02c4dedf95ea6642422b6fa121aa9172c949896fb9bdb38aa337

Observation 59abc877-3b58-4ef4-be89-dab5d51edd25 · outbound

This paper cites an unresolved cited work.

Fast Convergence of Policy Regret in Learning Stochastic Optimal Control Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-29T21:10:16.593076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T21:10:16.593076Z digest=sha256:bdd74a72651bb7351310caedb1fabb6c8f8b4f3581be754cb5b25d0514afc7a0

Pith citing papers

No inbound Pith citation observations are available.