Pith. sign in

Paper Citation Record · LEDGER

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs

As of 22 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2605.24345.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24345 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T15:10:49.492879Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact5
  • verified fuzzy42
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fce107f6-b1d3-43f8-bc1d-9c05dc460b1d · outbound

This paper cites Proceedings of the Thirty-First Conference on Uncertainty in Artificial Intelligence, 1--11.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the Thirty-First Conference on Uncertainty in Artificial Intelligence, 1--11

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.204120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:c79b6e9678420a503574d7a497d5569811fa583dcc97c033f07b5dd103188cc4

Observation 9a8142b0-a82e-487e-b742-f98aa026c379 · outbound

This paper cites Mathematics of Operations Research 48(1):363--392.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Mathematics of Operations Research 48(1):363--392

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.205919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:a396987bd245090eb2a6482cda1ace1c829839d536f7f690b00334e816651f24

Observation 21c821f7-3aed-4057-82e9-5c58e7c3d611 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning, volume 70 of Proceedings of Machine Learning Research, 263--272.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the 34th International Conference on Machine Learning, volume 70 of Proceedings of Machine Learning Research, 263--272

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.207655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:97bef863b5938d3144c6a6ad143569e2ddd7a13a06449d008213d2735d21daf9

Observation 4232c20d-876d-406e-8f9e-3d41ccb853e5 · outbound

This paper cites Proceedings of the 38th International Conference on Machine Learning, volume 139 of Proceedings of Machine Learning Research, 511--520.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the 38th International Conference on Machine Learning, volume 139 of Proceedings of Machine Learning Research, 511--520

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.233233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5cc337f9de4cad8694ca372983e708cb499cbc2e33ae2903e0053dc7b9761cd5

Observation abfb1525-1f8e-4219-a9f3-d25dfbadcbcc · outbound

This paper cites Proceedings of the Twenty-Fifth Conference on Uncertainty in Artificial Intelligence, 35--42 (AUAI Press).

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the Twenty-Fifth Conference on Uncertainty in Artificial Intelligence, 35--42 (AUAI Press)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.184949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5bb3fd74fef9335dea17a7de335f892ea17f5bff87ff958ce5c64280cc85c442

Observation dfbb9c4b-be19-4b46-b3d1-d060ee3514c4 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 36.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 36

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.148463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:cff2aeb6ec629cb65ab23390394213c451f0c470a3fa7ace21dbba580b1c29d6

Observation e40f4c44-11f2-4529-88ca-49b3e4ce1836 · outbound

This paper cites Electronic Journal of Statistics 3:114--148.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Electronic Journal of Statistics 3:114--148

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.146064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5cb7f86c8117f35b5524aa70a4b288eb0b64a5a486b4cbf8b0fe896ee583c15f

Observation 16dbf3b4-9f60-49d0-814f-534bb37079b2 · outbound

This paper cites Journal of Machine Learning Research 3:213--231.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Journal of Machine Learning Research 3:213--231

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.151451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5688d82f734fea5b6652730257e9e696c30e02256b2b8d628cb94005151089a7

Observation 1dc7f7c3-b265-4299-9d23-8948ec7dcc86 · outbound

This paper cites Operations Research 58(1):203--213, ://dx.doi.org/10.1287/opre.1080.0685.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Operations Research 58(1):203--213, ://dx.doi.org/10.1287/opre.1080.0685

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:14:46.208508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:c00d89cb3bc64712d08125b9d5df67b12ab3562521ed78ce9b409ce97bc6b42b

Observation dc26404d-ef28-403c-aa23-7b70082db56f · outbound

This paper cites an unresolved cited work.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-08T18:35:27.157493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:a6a21ff81ef7c8fc422598935712b01a0bfc1af2e4e06ecb025fed2036602632

Observation dbb3e075-3961-422d-a01b-c6990061fd55 · outbound

This paper cites Online Policy Optimization for Robust MDP.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Online Policy Optimization for Robust MDP

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:14:46.893972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:b1d60ae10bca4f4ac5d0dccc71283bfeff2cff072fed38c21da477e512df631c

Observation f36cfb98-ccd3-4bc8-b3e1-66059f69645a · outbound

This paper cites International Conference on Learning Representations.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs International Conference on Learning Representations

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.159214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:f9ee7be068bfd9e6a7abe17319afc8e02713330881296bf6391fc10c452cb1ee

Observation cae65240-fb93-4107-af57-7dd40386570d · outbound

This paper cites an unresolved cited work.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-07-08T18:35:27.219922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:08ddff2da23543a64d7d449dc05eaccf1d0e27280171285998b7457a06b17ac5

Observation 98f886db-c3c5-4c11-a32d-dce3a2414f50 · outbound

This paper cites Machine Learning 110(9):2419--2468.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Machine Learning 110(9):2419--2468

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.221961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:2dd95eed2e2f02d24e747d96ee83b513ccbf17d0a5a151abd866e4296ec019d4

Observation 4f607309-7731-4017-9425-c4ee617f33f9 · outbound

This paper cites Proceedings of the Twenty Third International Conference on Artificial Intelligence and Statistics, volume 108 of Proceedings of Machine Learning Research, 1431--1441.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the Twenty Third International Conference on Artificial Intelligence and Statistics, volume 108 of Proceedings of Machine Learning Research, 1431--1441

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.216093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5bd5b9ba4c602789c9010e79b42ef93a62ee9a12ede48c8bb382e69d56e6a942

Observation c49e8f8a-b28e-4ac9-8618-ac62c7270fc8 · outbound

This paper cites Foundations and Trends in Machine Learning 8(5--6):359--483.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Foundations and Trends in Machine Learning 8(5--6):359--483

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.213818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:857edb7917c7704100700a6c42fa0c502aa3244e406b3cb7921804089ec4ce26

Observation 72b766f7-6dca-4721-abae-c6ff742e48f2 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence 40(25):21278--21286.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the AAAI Conference on Artificial Intelligence 40(25):21278--21286

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.218107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:88f39f39114dec66a2c3c4cc746718abd902d34499615644557cd4f9ef60cbb6

Observation c59c3afd-cb5e-47e9-8b63-89cc045e9fe7 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 34, 22288--22300.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 34, 22288--22300

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.223967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:d498314b64fd343dbbe751c30b98ca1b36dcbc5a05af631c8f335b1dcb265e8d

Observation 15656899-95c3-4135-99aa-c1d10d313e87 · outbound

This paper cites Mathematics of Operations Research 30(2):257--280.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Mathematics of Operations Research 30(2):257--280

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.225674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:e1ce87f4a0f5bd263638d079c2b42c201bdcad245c54b480502be3cb4b60546f

Observation ad2f6106-2d51-492c-88cd-d5f373c08dd0 · outbound

This paper cites Journal of Machine Learning Research 11(51):1563--1600.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Journal of Machine Learning Research 11(51):1563--1600

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.227622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5403a5867aaeab9b3bba369b20916e40a4481837a19c11d9315500c47da6d074

Observation 5edd6949-e480-4fbc-bb1a-cd1e1cf6662b · outbound

This paper cites an unresolved cited work.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-07-08T18:35:27.211778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:2280484e5912f3d81283878926bc0379d4358afe40c180a1b3d69e1a4550825b

Observation 466e473f-369c-40ec-a233-b28e28d28148 · outbound

This paper cites Machine Learning 49(2--3):209--232.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Machine Learning 49(2--3):209--232

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.209586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:da08ba5ed85c1d8baafb40edc8f9a5a1ca521451da81056f16ebcb62b2ded24a

Observation 6d80bb4a-fd2e-415d-b207-5f98ea70f530 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence 39(27):28195--28203.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the AAAI Conference on Artificial Intelligence 39(27):28195--28203

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.239083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:60a1f18215a1152c125e472de9fe1f35d431e6e8b824a21a694659a2221b505a

Observation 2177d0ee-14fc-400b-9a4d-254c6a217d4f · outbound

This paper cites Advances in Neural Information Processing Systems, volume 35, 17430--17442.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 35, 17430--17442

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.192768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:17b8d425c8c1c43dfccbdc2b002308f2f624169d54ef1ab5371ca452957c04f6

Observation b4f5bd30-6044-4a5a-b70b-684ca731a3af · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence 39(25):26605--26613.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the AAAI Conference on Artificial Intelligence 39(25):26605--26613

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.186782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:baf7fc3688a89ce2fd2f6954206a60f8299b2d97c4b7f82c4f9a3aea2ad286e2

Observation d2075aaa-9d31-42b4-9491-0da5e1a9b915 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 37.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 37

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.188706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:57092007407629323c1e0c9d138dabd68579a1a2b7bacd7b6412e9c114df2c4f

Observation 29a596fd-4f49-4090-84ac-5621ace4d1d2 · outbound

This paper cites International Conference on Learning Representations.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs International Conference on Learning Representations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.190740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:2398eb692c84bd886e86217192fb2b75dcf1d06c74f86bc2c702cbea2b5e13a6

Observation 02306ffc-428f-40f4-b936-04f5d920300e · outbound

This paper cites Operations Research 53(5):780--798.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Operations Research 53(5):780--798

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.194414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:ebf1f517db36af80d0b54f6cd795052376f9056735a16d8bc3ed8063fb14dd51

Observation b0b80b24-1e6b-4a08-8c55-27dbc3711154 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 26.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 26

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.198834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:fa1a246281d594da6eaa2570bc89e5b17f3f37b1577b9b33dae3caff53d9a032

Observation b85bddfa-fc31-444e-b801-4e1e68e1de16 · outbound

This paper cites Posterior Sampling for Reinforcement Learning Without Episodes.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Posterior Sampling for Reinforcement Learning Without Episodes

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T15:14:46.903832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:fab9d3b4980e11c9b5d17dc598c4a9b52778700179504ba85556ba609a7ac9fd

Observation 48188fe6-b22d-400b-8044-9d11957c9a8c · outbound

This paper cites an unresolved cited work.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-07-08T18:35:27.179244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:296155907130158ff08558ec04aeebc43beb5eb4174b4a331813228b7acfcff2

Observation f8ede60a-79b3-4114-9ac3-6b3c2610301f · outbound

This paper cites Proceedings of The 25th International Conference on Artificial Intelligence and Statistics, volume 151 of Proceedings of Machine Learning Research, 9582--9602.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of The 25th International Conference on Artificial Intelligence and Statistics, volume 151 of Proceedings of Machine Learning Research, 9582--9602

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.182989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:a40ec41c6c567c31f4b3ec3db3bfac67de86c095cecdf18db7845ba999e393ba

Observation eb47cba3-2ab8-4c5e-bb08-276492fc0c74 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 35, 32211--32224.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 35, 32211--32224

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.181086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:d56a45d3ccef1d214ccc0798e238985c8d991a3cf7353ae4868181d664aaa658

Observation af0a346a-ae20-42e9-b31a-78f4a2a95068 · outbound

This paper cites Proceedings of the 23rd International Conference on Machine Learning, 697--704.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the 23rd International Conference on Machine Learning, 697--704

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.184763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:aebeeb004c8e0f592ab690c7759c44cff5e06a66ec0b5c95a275df096da40e69

Observation 3ade112a-8976-4fe3-9c07-16e80ca55dd9 · outbound

This paper cites an unresolved cited work.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-07-08T18:35:27.200442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:8cd8e2859427047af1f0e1adb030ebb7fe44fd55b480fb418aa46cc74cd833d0

Observation 33f1db7f-2183-437e-aba8-b90e092098ed · outbound

This paper cites Advances in Neural Information Processing Systems, volume 20, 1225--1232.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 20, 1225--1232

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.196353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:669898bec9161a84e83832c97510de38608df5f93cc7f75196f1c4603e712055

Observation f1e886c7-6c48-451c-b698-97a376048208 · outbound

This paper cites Foundations and Trends in Machine Learning 11(1):1--96.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Foundations and Trends in Machine Learning 11(1):1--96

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.237143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:e44c9aa1f8c7ad0db7699a5000ca3c8965115b4c58a7d017a6b9a7a3d2cd7071

Observation d786e6a7-f166-41b7-a275-45058984496f · outbound

This paper cites Proceedings of the 23rd International Conference on Machine Learning, 881--888 (ACM).

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the 23rd International Conference on Machine Learning, 881--888 (ACM)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.155654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:f88bb72952da209037e8a1565be94aed0ee0b81c712e91c6648d5577a87007ae

Observation b713729e-e110-441d-9332-f0a18d517a57 · outbound

This paper cites Journal of Computer and System Sciences 74(8):1309--1331.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Journal of Computer and System Sciences 74(8):1309--1331

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.161308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:7fc5571bf4c2c7cfdd9dacd1bdaebe1e99cff7252f2b5eec08c5967f125e1cd5

Observation 50c9e89b-ac18-4f82-96e7-67dc3d9857b4 · outbound

This paper cites Probabilistic Inference in Reinforcement Learning Done Right.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Probabilistic Inference in Reinforcement Learning Done Right

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:14:46.898979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5ffa19d12f90a4fcbca4e3361841fdb70e463cd473edc2a9bcb1a965080b92ed

Observation 1c4c0c84-94ae-4569-b292-c1a705cec25a · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning, volume 162 of Proceedings of Machine Learning Research, 21380--21431.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the 39th International Conference on Machine Learning, volume 162 of Proceedings of Machine Learning Research, 21380--21431

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.229521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5c623b8d9bb84342d6ed96df184a9a840e0997535c34bb1971a419db64d7066f

Observation 243e2625-4cf5-4ec4-bc57-29d8b4e61517 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 36.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 36

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.235077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:6343ff650ae176390fdae50a7d70b944504919b6f2a1e8b20767c49010f58c0f

Observation 292d6809-e8e4-4adc-a5bf-22d81513b3e8 · outbound

This paper cites arXiv preprint arXiv:2509.14077.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs arXiv preprint arXiv:2509.14077

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:14:46.896400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:6bfae946ec103d87b642dbb0942e0665fc63ab0a5c0c133b7bd010ce7edef3f6

Observation 2b2ef86c-c21a-4cc1-8a1b-106b3ba4d0f9 · outbound

This paper cites Advances in Neural Information Processing Systems, volume 34, 7193--7206.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 34, 7193--7206

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.231366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:80ff7e4801217b63f1153bce2778d212a9c8543088269c987d68aa9a2ec33f45

Observation 3827dd32-194e-463f-8321-3b593b621327 · outbound

This paper cites Mathematics of Operations Research 38(1):153--183.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Mathematics of Operations Research 38(1):153--183

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.219737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:48e73680be66cf19bf2e89539f9b99b05b60b755e3cff3802e4ad94c9a4d21b2

Observation 9edd5da4-363d-4b0b-8da6-9db388c2b2d2 · outbound

This paper cites SIAM Journal on Optimization 28(2):1588--1612.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs SIAM Journal on Optimization 28(2):1588--1612

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.215913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:6a941575eb76e5743b22f6e3b988e81f678d77a7da1a0da60622c20e0cccbf90

Observation 15cd185a-ffa1-4ea7-b4dd-80e8d0e5102a · outbound

This paper cites Advances in Neural Information Processing Systems, volume 23.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Advances in Neural Information Processing Systems, volume 23

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.197905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:5e5d46dcbb90b7745830be1f3fff383927e115ce7c9458f5f742147045875add

Observation 257d9bae-8bf7-4348-a294-7b3bf0627b6a · outbound

This paper cites Reinforcement Learning Conference (RLC).

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Reinforcement Learning Conference (RLC)

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.200168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:ea5c096d32b70d2417a744181c964124ab5adc50e941982bf1275a87ac0ddddf

Observation 4c2a74bc-163c-49c2-9a55-a1d03bca2b20 · outbound

This paper cites Safe and Robust Reinforcement Learning: Principles and Practice.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Safe and Robust Reinforcement Learning: Principles and Practice

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:14:46.901529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:770d225455dbe5e8f402b166f3ba8c976cb6d7db60040372b2a13b65dd60f47e

Observation fdd82cc3-a609-477d-a34d-c9c3700968ae · outbound

This paper cites Proceedings of the 2015 Winter Simulation Conference, 3714--3724.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of the 2015 Winter Simulation Conference, 3714--3724

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.195449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:8e844c7394bd654452066cedc8d6d0bec60c3035a3ed0cd15a3bb0fe4df7ef79

Observation 90e36eac-1004-4c62-9637-94590eba0627 · outbound

This paper cites Proceedings of The 24th International Conference on Artificial Intelligence and Statistics, volume 130 of Proceedings of Machine Learning Research, 3331--3339.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs Proceedings of The 24th International Conference on Artificial Intelligence and Statistics, volume 130 of Proceedings of Machine Learning Research, 3331--3339

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.177103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:76425c0f4fc7b3a3ab39cc987af801da0cc4b5f9386ffd0e8662777bdfb19457

Observation 1567941c-2554-45f2-89a1-8aac3a5ae272 · outbound

This paper cites " * write output.state after.block = add.period write newline.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs " * write output.state after.block = add.period write newline

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.202322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:cf79029247e28b0f5bec432ba96c8b5eb758baccebbb20b23542e28483a63b54

Observation 43d3408c-a02f-4d20-9162-792547b2f015 · outbound

This paper cites write newline.

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs write newline

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T18:35:27.163823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T15:10:49.492879Z digest=sha256:09d3348691a6cdc06ecedaddc4a877498d2c130f4557e2370b75054b4df8aeb1

Pith citing papers

No inbound Pith citation observations are available.