Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T23:51:47.193267Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 90 of 90 outbound references and 1 inbound Pith citation observation for arXiv:2606.00151.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T23:51:47.193267Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T00:41:12.548554Z
A source-named dated measurement, never combined with another source.
Source: cited_works
90 of 90 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2f396977-e942-4c10-a86e-1184be9f322a · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying and Barto, Andrew G
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66aa3f75-3006-47aa-a333-3a36e9217675 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Why generalization in
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04384a00-28f4-467f-97f1-13ca01c50d49 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53c2b838-1da1-4cbe-b1b7-698de14a6aff · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88d319e0-6419-46b4-9c04-9e97da1d74ef · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Strehl and Michael L
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1a75296-f3b7-4d3c-969a-4a1c8f41af11 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying and Li, Lihong and Wiewiora, Eric and Langford, John and Littman, Michael L
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cac9b50f-16f8-4f48-b2e8-79931f21a1b3 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a07c8584-9b9f-4232-bdd2-edf75d4e10bb · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying and Littman, Michael L
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59aff438-2a21-48df-9140-6e78e23b3113 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1898857-5169-4006-8d7f-f43f85ffee79 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6d2efc7-f9e1-4c26-8b9a-95bb7e50b997 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Stadie and Sergey Levine and Pieter Abbeel , year = 2015, journal =
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f742a8ed-324e-4f92-991b-4621a0d5ba31 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying and Darrell, Trevor , year = 2017, booktitle =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aadd9723-2509-4f22-b7c2-0a3a556546d5 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Thrun and Knut Möller , year = 1991, institution =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee5ab933-a037-4513-9aa9-7314366988a5 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Gomez and J
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9174b7f-74b8-497c-ba17-464e9e84c7ac · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying IEEE Transactions on Autonomous Mental Development , volume = 2, number = 3, pages =
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 260a3bf3-24f5-4659-a5f2-adadb7e503f8 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15f6e43b-1e0c-437a-aa02-274ca91c8b7d · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2a274f8-109e-4918-b5e7-e00223366ff1 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b18caf4-3be9-4d06-82e7-872744a98fad · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a89c7a21-e6ef-468d-bb3a-d4a49f9f35e4 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bce3a46-e364-462e-9080-4b1e623d6d3b · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b4f4011-ef96-42be-9425-d3cf164cb876 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86e05baa-62e5-4ccc-baf5-95732d665c2f · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Nature , volume=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a93fc4a4-ba43-46d4-90d1-3b61193c6348 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Williams and Jing Peng , year = 1991, journal =
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8c38882-5a73-40eb-8231-ec2e54277c87 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Machine Learning , volume=
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f6ec25-bb09-4f70-9763-9eee5a33c29a · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b423f32-d9db-4b41-81a5-6244e31b96ba · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 945c1c00-efde-4466-b982-b7ca67106a54 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12e31e08-fe0f-4649-999e-a1df50a6c2e0 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 204384a4-640b-4a96-8e57-f60de5cc534d · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d490338-b480-4b78-8606-e7bf3ac08c10 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb4fbd68-dfe6-4475-9fe9-c34b3080ec08 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying and Maas, Andrew and Bagnell, J
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41f463b7-0ea3-4376-8435-6a0976766924 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c9b177-90b0-415f-a706-9c1ea5596376 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc57d2cd-0c27-4159-8e10-8f6b87a00c12 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Machine Learning , pages=
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8037c2df-0873-45ba-88e9-13d5aa173e85 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Deep exploration via bootstrapped
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfe5d583-bb38-4acb-8b69-969becd2d622 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 790c50b6-f7f8-459f-ab4e-83884c727cf3 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Journal of Machine Learning Research , volume = 20, number = 124, pages =
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e8f8a2e-9272-462f-be9f-6d8a5cff5e00 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Efficient exploration through bayesian deep
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 438fc4de-93f1-490d-bf94-d21145c9f9de · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , volume=
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b134b419-7dad-40cb-bf66-5df1a5d7ad46 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Uncertainty in Artificial Intelligence , pages =
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e84a7a12-3701-4ae3-8653-c9a0fb5224da · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Machine Learning , pages =
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76bf6ead-c15f-4fdc-a12b-5564daebf609 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , volume=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9ca94e6-ffd6-42ee-a980-4a2842f250ca · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Machine Learning , pages =
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1eea2419-29fd-4511-bceb-91dbba4b73fe · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Machine Learning , pages =
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44906199-4c35-4447-8ba6-2653663bb128 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , year=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6feed398-e68a-4141-bcee-40437d1841de · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8971fd8-68e1-40f2-aac3-7acd49148511 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume = 36, pages =
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3c6aa87-14af-4f49-a216-95e89642a080 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , volume=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4770aa04-f0dc-4cc7-addc-5827a65ec306 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebeb77cc-a45a-4db5-bd7d-08656538a854 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Journal of Artificial Intelligence Research , volume = 47, pages =
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f35cec72-c1ac-4dc0-8058-8c0cb3f419c1 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume = 35, pages =
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9376ef5c-bbed-4566-9521-66e8dd21aa8f · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume = 34, pages =
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ed91220-3dcf-472c-8d1e-47123858796a · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Journal of the ACM (JACM) , publisher =
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ef9e9fe-2029-4303-ad05-3fce038d6680 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Artificial Intelligence and Statistics , pages =
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99f1bc58-71f8-4716-be71-3f0adaf9deec · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Biometrika , publisher =
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 693e94f0-acb1-405d-90bf-c6a51886ff81 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Machine Learning , publisher =
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c5ec34c-6783-45e4-8061-4233cc7c5622 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Journal of Global Optimization , publisher =
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 185c9c26-f80a-4d18-88a8-7f0be66c25ff · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Proceedings of the 19th international conference on autonomous agents and multiagent systems , pages=
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f4c9f12-2010-46ba-ab38-3c65f90785ea · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9caadd62-01d0-47a6-a79d-b5c3e8e573e6 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b25580dd-a8a4-4149-8471-00528ee90289 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d75a9694-8abf-4779-b6c0-734acef5c68f · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Uncertainty-based offline reinforcement learning with diversified
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3df2e2c7-6eae-4595-a26f-20e7a3ecd917 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Maximum entropy
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 572f28a8-d80b-4738-8848-00c15138f6b6 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume = 33, pages =
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9697fb8-6d9b-40df-a1c1-fa2ec88aeba8 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ff0a4c3-26f9-40c1-945c-0ec364e656ed · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , year=
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7633eec-1e77-4b2a-82a6-a345f558e068 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a57cbaf3-586b-4095-9bc6-b8f2ce585988 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume=
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f91ad50-d475-4ac4-903e-ea3e0412b2c3 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying No representation, no trust: connecting representation, collapse, and trust issues in
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 270b1b7d-77b5-41d1-9369-acd4106d393d · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Proceedings of the 41st International Conference on Machine Learning , pages=
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2eab59b-02fc-49f6-a0bd-595c52963911 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , year=
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e852d4ca-744b-4c80-99e0-1079980139a5 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cea13e32-4892-4b0b-8283-d59626da5d5e · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Araújo , title =
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97b41722-d851-4be0-ba92-03f2985d1941 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc9cc110-affc-4a1b-b501-d8ca5638eac7 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , volume=
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c95f9b1a-c6e5-4947-8b2e-788a638d29eb · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Machine Learning , pages=
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcecc30f-967a-46a8-a6b5-0c3ed007f485 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying 2022 , url=
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e70d6bb2-5768-4c06-8a2a-962d293d7981 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Reinforcement Learning Conference , year=
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3cdab19-2a70-4cc9-a565-aba07326e981 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Learning Representations , year=
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3848e43d-653c-47a7-b0d2-db3e19e7c5eb · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Beyond Softmax and entropy: convergence rates of policy gradients with
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f42f263e-72b7-456d-b310-839281a78f17 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying 2024 , organization=
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9360dd9c-942a-4c88-bd0d-325c619d7363 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume=
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 644b7d82-c7aa-4d19-a5a0-708845348a40 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Management science , volume=
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc88741a-c101-4a7c-bbf9-652670b5d889 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Advances in Neural Information Processing Systems , volume=
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e8c564d-0386-4053-aff1-a869081b1abf · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence,
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e76756f-41eb-48fa-ac7e-06ae6f89fd51 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Journal of Risk , volume=
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0265642c-0e7b-44d2-a0da-edae386085ec · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying Unresolved cited work
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ded8b7-c7ab-4216-b153-b5a2126a6979 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying 2023 , publisher=
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ba91134-924e-42c0-8610-5886b24dc0e0 · outbound
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying International Conference on Machine Learning , pages=
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f06b6f3-348d-4473-9602-e638dd50a9ab · inbound
Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.