Pith. sign in

Paper Citation Record · LEDGER

Efficient Skill Discovery via Regret-Aware Optimization

As of 9 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 0 inbound Pith citation observations for arXiv:2506.21044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21044 v1

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:41:38.230312Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

74 of 74 outbound references displayed

  • verified exact2
  • verified fuzzy51
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 90c85963-0971-41b0-9aa5-55b627618389 · outbound

This paper cites write newline.

Efficient Skill Discovery via Regret-Aware Optimization write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.145979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.145979Z digest=sha256:3e7f05682d5322be76cc30f3b1e6bb0e401774fd501e9a9122db8f7e4f1cb207

Observation 1da8fa44-ed1a-4839-883e-e382ad82a665 · outbound

This paper cites M., Crump, T., and Far, B.

Efficient Skill Discovery via Regret-Aware Optimization M., Crump, T., and Far, B

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.322974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.201982Z digest=sha256:b2f0f21882563ce95466e01e5b311264dec171a2521e888317889da381cbe5fc

Observation b46a063f-888c-427f-b844-28504c636e07 · outbound

This paper cites Hindsight Experience Replay.

Efficient Skill Discovery via Regret-Aware Optimization Hindsight Experience Replay

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.306850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.306850Z digest=sha256:d4011d874e449c20683495c54bcc812ff0145dc66c9d753c7af8a43fd3031c3d

Observation 88f1255f-08d8-45f4-a571-a72faee19108 · outbound

This paper cites TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations.

Efficient Skill Discovery via Regret-Aware Optimization TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.424348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.424348Z digest=sha256:acd51ed35ac0657c79bcbcfd80e71022fdb75967c875cbf371ecb99bb389a7f6

Observation 1ea52e5c-7a75-40e5-8676-473c33a22ed3 · outbound

This paper cites K., and Konidaris, G.

Efficient Skill Discovery via Regret-Aware Optimization K., and Konidaris, G

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.310533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.518360Z digest=sha256:8718c84c3c45941ff4be8f61471f57b8a15dac6718384207fdec28c4c76c11d0

Observation 0e414a57-7eaf-4db7-9a17-26479fb26d17 · outbound

This paper cites Constrained ensemble exploration for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Constrained ensemble exploration for unsupervised skill discovery

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.297553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.609290Z digest=sha256:1bb24704a14620bfeb8125a9346fb382e8ae33b49dd7c3c99b8287aa48d3d1b7

Observation 9a4094e2-9f68-4cd4-bdfe-ea633f5bcc5d · outbound

This paper cites X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U.

Efficient Skill Discovery via Regret-Aware Optimization X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.283764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.703614Z digest=sha256:bb73d42c4e3b48f4982e9e0e5ef57de208cc36777345feac6e4a7d872d80c839

Observation 0f826a26-1c03-4767-ade4-ceeb660bfa10 · outbound

This paper cites Exploration by random network distillation.

Efficient Skill Discovery via Regret-Aware Optimization Exploration by random network distillation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.270252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.785283Z digest=sha256:7b45fdcb48a43040e7c71ddc8c47651d4444546d06a890cf2ec729f3f7b815d9

Observation 55067e30-d11b-4a7b-9452-07042e042327 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills.

Efficient Skill Discovery via Regret-Aware Optimization Explore, discover and learn: Unsupervised discovery of state-covering skills

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.256300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.836290Z digest=sha256:796a558605323aaacb8e23e7b46b22607d3c0570e8d69d11d159ff45f175f58f

Observation 6f836df7-b49c-4975-87b2-7d70de188e82 · outbound

This paper cites Language as a cognitive tool to imagine goals in curiosity driven exploration.

Efficient Skill Discovery via Regret-Aware Optimization Language as a cognitive tool to imagine goals in curiosity driven exploration

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.241660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.917944Z digest=sha256:9b2b16f0ee0dc829f98dd2339f7ec1a202f380de92c840ffd8aa94acc56672ff

Observation 0b03d062-5f28-41ad-bee5-78a4176256fd · outbound

This paper cites Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey.

Efficient Skill Discovery via Regret-Aware Optimization Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.227451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.005894Z digest=sha256:14fce1f82f5b6e865728f24b78890d167d2f87a71ff4a222989df224a37f1f9b

Observation ba3fe812-2a9d-4457-abdc-473cae2bb186 · outbound

This paper cites Goal-conditioned imitation learning.

Efficient Skill Discovery via Regret-Aware Optimization Goal-conditioned imitation learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.212302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.090668Z digest=sha256:e5bc08a021266d36dbe1d5edf3174bb52cd53b613050620e3382c2a171f17fc9

Observation a5494acd-cb81-4cd8-a6e5-f246fab0f8bd · outbound

This paper cites Adversarial intrinsic motivation for reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Adversarial intrinsic motivation for reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.198143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.203128Z digest=sha256:e3257ec99c76832e650be2d0928c70ae140a62b8829a718616ed1ff6af5e60b7

Observation e284d72f-c053-4426-b2b3-0abe976a1ce9 · outbound

This paper cites Diversity is all you need: Learning skills without a reward function.

Efficient Skill Discovery via Regret-Aware Optimization Diversity is all you need: Learning skills without a reward function

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.184132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.247740Z digest=sha256:1c70857aa0b6d25a844302504165de219010e1a63c16158ea654c32aa3402285

Observation 143187d5-3f4d-4258-9ef2-695f91139aa3 · outbound

This paper cites C-Learning: Learning to Achieve Goals via Recursive Classification.

Efficient Skill Discovery via Regret-Aware Optimization C-Learning: Learning to Achieve Goals via Recursive Classification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.332286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.332286Z digest=sha256:7b009d5eb7c89daa6cfe48d481fd0ffbaa80b3ea2843f4b1876d0d33449e5e01

Observation 20c0ea5b-129c-4409-bda9-7de5ec4a561b · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.170092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.417950Z digest=sha256:40d3f83eac9605d9c6017b12a79b7e5206e63a8b8b0a65446ec4ec61d2e8d423

Observation d9147cd4-2c24-4fa6-b982-d8eec0274326 · outbound

This paper cites Curriculum-guided hindsight experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Curriculum-guided hindsight experience replay

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.156692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.468891Z digest=sha256:293fc329dd9a163bf896165cd723c07c1ab084d9098e7333c186504550a5f90e

Observation a5e18089-2bb6-4ebf-b2dd-aa137296a80a · outbound

This paper cites Automatic goal generation for reinforcement learning agents.

Efficient Skill Discovery via Regret-Aware Optimization Automatic goal generation for reinforcement learning agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.473619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.473619Z digest=sha256:bca5f22d309ab0679c292a1364c8e5dbc51cc451d7b4e5721bc580ba5d95b6ed

Observation 7d6cf578-631c-4d71-849a-32e0cf78ae83 · outbound

This paper cites Accuracy-based Curriculum Learning in Deep Reinforcement Learning.

Efficient Skill Discovery via Regret-Aware Optimization Accuracy-based Curriculum Learning in Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.578090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.578090Z digest=sha256:da3cde4b4dec9cee7bb0ad0902f49e9eb56ea3f618be14f7334420f51fab127f

Observation ccb63044-06ba-4f5f-8633-324719177297 · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2020.

Efficient Skill Discovery via Regret-Aware Optimization D4rl: Datasets for deep data-driven reinforcement learning, 2020

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.133429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.762057Z digest=sha256:2072ff45cee01a7cf0b13494973335cf34fb0bf1252c66d6db0295029877e0a9

Observation 77335ca6-e087-4292-8657-b2892f585840 · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Efficient Skill Discovery via Regret-Aware Optimization Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.963486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.963486Z digest=sha256:646853fe0f74f793ad9247460738a35ec02e3b018e0a29a3b1075e3383356bdd

Observation 53a1012f-57ab-4f05-a980-b04548088400 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.121608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.121608Z digest=sha256:f348a92bd8558ed8d0e3964b7070b6a25a2c37bc76c6c5a9ce960b6a2fbdfe85

Observation 50abf721-f062-4549-8f28-ca8c2852b4eb · outbound

This paper cites J., and Wierstra, D.

Efficient Skill Discovery via Regret-Aware Optimization J., and Wierstra, D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.110596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.261687Z digest=sha256:ea4e2f0f587b621651c0eba4195b44724d29824bcd0e363befbcccfd30acb381

Observation 37668b98-4ab1-4136-be9f-db203f06ca2b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Efficient Skill Discovery via Regret-Aware Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.471301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.471301Z digest=sha256:12da9976a6e9d20f5ac6bb2bfdf7d483ebe092e6bcb740f3903c201b9fd829a5

Observation a1cbfc2f-458d-472e-a9bb-20c95be91d63 · outbound

This paper cites Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation.

Efficient Skill Discovery via Regret-Aware Optimization Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.461935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.692807Z digest=sha256:bebfbc8d658a20ab897f02b420778a0fc1429a8e456c86f2883134fbaae74cc0

Observation 33de2148-99d3-48aa-8870-312b786ce215 · outbound

This paper cites Exploration in deep reinforcement learning: From single-agent to multiagent domain.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: From single-agent to multiagent domain

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.089016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.858079Z digest=sha256:84972eaf945cf10353e658d44d2f6a65014bd792ca2878aeff7342786a8ea988

Observation 3a774f42-b30f-4306-8538-3989effc3200 · outbound

This paper cites Open-Endedness is Essential for Artificial Superhuman Intelligence.

Efficient Skill Discovery via Regret-Aware Optimization Open-Endedness is Essential for Artificial Superhuman Intelligence

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.029860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.029860Z digest=sha256:485b6dc4eccada4093ac4e8cd11d901393bf2836b08c82485371e2ae2050c0f2

Observation ce161550-23c4-49d2-a496-0089d72ad128 · outbound

This paper cites Unsupervised curricula for visual meta-reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised curricula for visual meta-reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.074905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.228723Z digest=sha256:1322fe10ebd584d4909641ba7368e3d0065616181d0dd730d30bded55bfbf741

Observation 63d64998-cf2b-468f-9fe6-c4c89b7cb46e · outbound

This paper cites A comprehensive survey on self-interpretable neural networks.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey on self-interpretable neural networks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.398189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.398189Z digest=sha256:e9fb448cb2cf8bb810961c8ad6acb4b9f1f2beecb8675995d1a5b79eb27c6376

Observation f843fe81-8fb2-4a69-a52a-25fa443fe92a · outbound

This paper cites Replay-Guided Adversarial Environment Design.

Efficient Skill Discovery via Regret-Aware Optimization Replay-Guided Adversarial Environment Design

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:41:38.328725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.601070Z digest=sha256:0ae2fbaefbe0974d780d15348b94550f5f50451af58d11a3912e14cf2c3e1061

Observation 1fabd924-f58b-4583-b50b-12a9223ad4f0 · outbound

This paper cites Prioritized level replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized level replay

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.061604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.669110Z digest=sha256:c02c617caf0b9778bcaeac05fe141c90e4f8e19ce59f66f71344c24ac40a3512

Observation d4f15ff1-9b76-4682-98e1-2f2825d64cbe · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.047719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.714340Z digest=sha256:5f470d484a20ae5a092ba803aee4a65b6b119e51e21c60ab1a877929a5ac77ad

Observation 844b097e-fc4a-44d9-959f-3033e3aa5941 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.740268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.740268Z digest=sha256:c03b4158922cf77ecece9b55088a181d74b913cabf50e0a8591e211a58a7d7c3

Observation 3b147d4b-9840-417d-a8ce-3fcd0f63d7b5 · outbound

This paper cites K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J.

Efficient Skill Discovery via Regret-Aware Optimization K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.026333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.814734Z digest=sha256:fbdc6fc81c35107991446f8462e9f847d2248d5f95d2487d8627fcb7a62d991c

Observation e96c04cd-b09c-4589-87e6-e8d3b65bec21 · outbound

This paper cites Unsupervised skill discovery with bottleneck option learning, 2021.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised skill discovery with bottleneck option learning, 2021

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.012568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.865430Z digest=sha256:053611b1d4505daebb1fe88aaca10f65b58e024774e40aef28b4abc271fa8189

Observation 31c84aeb-d9aa-468b-842f-3722166e7392 · outbound

This paper cites Active world model learning with progress curiosity.

Efficient Skill Discovery via Regret-Aware Optimization Active world model learning with progress curiosity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.998539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.942087Z digest=sha256:7f4838fe8854b5af9039b8c8cb53efb1fe2f37cd5914088cbafd30c0c5f1927c

Observation 661f3f1b-5f3b-4fb8-947c-aef0e73963a1 · outbound

This paper cites Exploration in deep reinforcement learning: A survey.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: A survey

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.985466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.015977Z digest=sha256:2f50d00b723ec0ac7c0157a5769e3cd81580c7c7a472dcbc7fd39dd6d5f82b5e

Observation b6d116a6-a48f-4cab-8144-738129961de2 · outbound

This paper cites B., Yarats, D., Rajeswaran, A., and Abbeel, P.

Efficient Skill Discovery via Regret-Aware Optimization B., Yarats, D., Rajeswaran, A., and Abbeel, P

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.971539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.079101Z digest=sha256:d67fa0694b12e0e623382661751b30d99c5c9e1233e604e524ff151fbc456681

Observation 7fb3eaea-27e2-4d7f-b4c3-a3b8e91864d3 · outbound

This paper cites and Seo, S.-W.

Efficient Skill Discovery via Regret-Aware Optimization and Seo, S.-W

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.957700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.143371Z digest=sha256:bd90f55a40f6ae5a251d6058cea81b7868e612de051a7e0d71d52c2fe52a804c

Observation 95fd50be-07ad-4499-b263-8839f0ba694b · outbound

This paper cites Revisiting graph adversarial attack and defense from a data distribution perspective.

Efficient Skill Discovery via Regret-Aware Optimization Revisiting graph adversarial attack and defense from a data distribution perspective

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.944068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.200699Z digest=sha256:0a6c82d89c1cb946eeb215e78bd210f14f1b839255e4e1740f57b9bf344cdffb

Observation 5cbd798b-7508-4414-98dd-1f02158d4e1f · outbound

This paper cites Boosting the adversarial robustness of graph neural networks: An ood perspective.

Efficient Skill Discovery via Regret-Aware Optimization Boosting the adversarial robustness of graph neural networks: An ood perspective

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.929948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.273877Z digest=sha256:8becec074b9f2853dba4e582d18b6a31bb3df61293a8fe17400cae340b105a78

Observation 8972b79e-4f56-4c4d-879d-e3ff63f6d453 · outbound

This paper cites A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals.

Efficient Skill Discovery via Regret-Aware Optimization A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:37.321403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:37.321403Z digest=sha256:15c3d997ea32b423ab3320af433add1d1dbdd6e15f0fa1e810cc6db5dace6a48

Observation 17c92cf3-8124-4217-8929-958f476c3f8b · outbound

This paper cites Choreographer: Learning and adapting skills in imagination.

Efficient Skill Discovery via Regret-Aware Optimization Choreographer: Learning and adapting skills in imagination

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.916681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.365884Z digest=sha256:b7356f67384530b4fb39f46804f198b48288e1789cddba4c191dc3e855ef65f0

Observation 57d2ebd6-1b94-4784-b061-afed5604589b · outbound

This paper cites Planning with goal-conditioned policies.

Efficient Skill Discovery via Regret-Aware Optimization Planning with goal-conditioned policies

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.901968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.462003Z digest=sha256:adc2832135f7212175953b0898fcc9df260545eea7161857eaf4ed4353103b52

Observation 06538529-6818-4526-94fe-164caa8d0a93 · outbound

This paper cites Wasserstein dependency measure for representation learning.

Efficient Skill Discovery via Regret-Aware Optimization Wasserstein dependency measure for representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.889465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.499235Z digest=sha256:32fb1395dce044e193868379eb13c1ab0aaa851acae0738db230c5567efa7246

Observation cb1c474e-9636-4702-886b-d7f66500ee3f · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Lipschitz-constrained unsupervised skill discovery

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.876381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.582284Z digest=sha256:d3c48d2c1c91bea21bebdb46cd50363aef66af327f1cddb6dd336e811df0559c

Observation 2318ca52-4a84-4898-b339-788fcc50afb8 · outbound

This paper cites Hiql: Offline goal-conditioned rl with latent states as actions.

Efficient Skill Discovery via Regret-Aware Optimization Hiql: Offline goal-conditioned rl with latent states as actions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.863376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.652587Z digest=sha256:39be9af9e2c88278c0fb62759a5896f8a9e23bfd26ccc3ec04a683bb72b517d5

Observation edf572e2-741e-428e-8fa2-e7cf5223703b · outbound

This paper cites Foundation policies with hilbert representations.

Efficient Skill Discovery via Regret-Aware Optimization Foundation policies with hilbert representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.849816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.697677Z digest=sha256:7f321c0b3ab90b04b6ea9930883ec4326fc662fa74e968c579959ee5fb25a8b9

Observation 1c666100-dff7-49e6-9bb4-4666000b500c · outbound

This paper cites METRA : Scalable unsupervised RL with metric-aware abstraction.

Efficient Skill Discovery via Regret-Aware Optimization METRA : Scalable unsupervised RL with metric-aware abstraction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.837059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.797162Z digest=sha256:c6c3297a6e678e998d63669840bd77726f0f72a0969f183e00d37948c8cd70ea

Observation b04ca4b0-c24e-47a8-a6e3-beb4780880fa · outbound

This paper cites Evolving curricula with regret-based environment design.

Efficient Skill Discovery via Regret-Aware Optimization Evolving curricula with regret-based environment design

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.824248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.867902Z digest=sha256:80f2304405bbe90c7dae628ec2973988e4d108972c08c340cb18dea201bac5c2

Observation 06fc72e3-eb0c-41b2-aeec-450f54d74c22 · outbound

This paper cites A., and Darrell, T.

Efficient Skill Discovery via Regret-Aware Optimization A., and Darrell, T

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.810397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.943248Z digest=sha256:513a47114e419508fa08fa19e044d0dca5a4f761a63de6e732bf30f9f60be2cf

Observation 99fba55d-032c-4e6b-9bbb-6e0f2bb19b85 · outbound

This paper cites H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S.

Efficient Skill Discovery via Regret-Aware Optimization H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.796256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.036854Z digest=sha256:b2df380a44d264569d9b262c5054a98b4bc29c3c8c609b7cc0d4dde2df27c0a6

Observation 9cfa1ce2-4d30-42d0-8d0b-d85150d118b1 · outbound

This paper cites Automatic Curriculum Learning For Deep RL: A Short Survey.

Efficient Skill Discovery via Regret-Aware Optimization Automatic Curriculum Learning For Deep RL: A Short Survey

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.079354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.079354Z digest=sha256:3ccefd7511f2b8b3a40bcc56b1b5065a9906a9a841cc0e2aa47c57a352dd5f50

Observation 428819a6-0947-4e3d-90bb-107c5f699bdb · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:38.783134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.121688Z digest=sha256:5c9e75d0748c6cd6661fcb0c8bb40495dd9a2db9c184bad018d2f13d59ab9892

Observation 20125f29-6832-4238-89d4-625fbd6c8abc · outbound

This paper cites Prioritized experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized experience replay

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.770421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.131699Z digest=sha256:8afce0dc00cb9a921c237b319b4b8fd1b1318a3622f5abd222e57988c47c0a58

Observation c49c44bb-ec1b-4384-93b5-5b9283d4a7b8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Proximal Policy Optimization Algorithms

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.139133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.139133Z digest=sha256:5c06486aa4309318cb9661eb08e78a3b6e4247f023337c7f4c36e6ba65762bbe

Observation 37147148-54ea-48b5-ae05-9d0211ef2774 · outbound

This paper cites Dynamics-aware unsupervised discovery of skills.

Efficient Skill Discovery via Regret-Aware Optimization Dynamics-aware unsupervised discovery of skills

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.755575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.149622Z digest=sha256:d12a0885bc23b0cd7ff7154c5cae0060ecbc8ef513a3edf1014922f6b7835bf7

Observation 3ece75c2-8146-4a87-aac1-db7807dc5998 · outbound

This paper cites Deterministic policy gradient algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Deterministic policy gradient algorithms

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.740517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.157683Z digest=sha256:1b5356926ff1e6410da29b9174cf6c4dd20806c97bfdf3592ee3915bf4fc7550

Observation ae196208-c010-45c4-9dee-eea8293b97e0 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play.

Efficient Skill Discovery via Regret-Aware Optimization A general reinforcement learning algorithm that masters chess, shogi, and go through self-play

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.725250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.165986Z digest=sha256:f501748d8bd57e16f8055320c63765d0ee25180d41ef1805ac6e7d9b5bb07f37

Observation 05395d97-732b-444f-971b-fd92e2885772 · outbound

This paper cites Intrinsic motivation and automatic curricula via asymmetric self-play.

Efficient Skill Discovery via Regret-Aware Optimization Intrinsic motivation and automatic curricula via asymmetric self-play

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.710934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.170823Z digest=sha256:1c4cc88bce51a8d42ec6e4b3b4f113cb57e02825784e599ffb08895c31b727b8

Observation 72c02ece-e109-43a7-a90a-62125a6fec8f · outbound

This paper cites Policy continuation with hindsight inverse dynamics.

Efficient Skill Discovery via Regret-Aware Optimization Policy continuation with hindsight inverse dynamics

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.696597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.174661Z digest=sha256:7596fe3ad4f4beecfd6739189db01e99a05d6535ce37f97364bc8403cd1d536a

Observation 9343b173-20df-48ed-9c25-1f9d65641445 · outbound

This paper cites Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections.

Efficient Skill Discovery via Regret-Aware Optimization Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.682455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.179067Z digest=sha256:7ad95983e1fb009b47f6c05609a0d94bd670f09253e5823ff78babbadc3c9097

Observation fdad43a7-eb70-470d-baff-90cdd4c02170 · outbound

This paper cites Market-aware long-term job skill recommendation with explainable deep reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Market-aware long-term job skill recommendation with explainable deep reinforcement learning

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.667895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.183587Z digest=sha256:1e3c58a687ecfb2a464c18f59b5747fc5a6d6cec6deb728eb9e11deec39bfa26

Observation 8abb6957-0f84-4e6a-b131-b2d2787a6ba8 · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.188231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.188231Z digest=sha256:f0ebf1d3f430fcf13cc856741afa026e69c572e4d10fb43e29db284b29757324

Observation 2c1ae144-c4ff-4eb7-bd6d-223d2bb13244 · outbound

This paper cites S., McAllester, D., Singh, S., and Mansour, Y.

Efficient Skill Discovery via Regret-Aware Optimization S., McAllester, D., Singh, S., and Mansour, Y

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.191955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.191955Z digest=sha256:c3bf7bbe7214510ebdd8886e708babc25bb101d02eb1c3f1a633f78a60ccad72

Observation 0c85fde9-bacb-4a1b-8a53-ba00562a2777 · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Efficient Skill Discovery via Regret-Aware Optimization dm\_control: Software and tasks for continuous control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.637023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.196846Z digest=sha256:70148850fee04c7238121fefe719220dd0b1ed94488b01f65da4d6868113c7aa

Observation 7fe00edc-7e18-485c-8903-9a3394f52bc8 · outbound

This paper cites M., Mathieu, M., Dudzik, A., Chung, J., Choi, D.

Efficient Skill Discovery via Regret-Aware Optimization M., Mathieu, M., Dudzik, A., Chung, J., Choi, D

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.623333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.201024Z digest=sha256:ef956ee7c92411bce58fce3b6e392591df427a8d110c1a19925ccd3b5d274347

Observation d7b4d0c5-2d5f-4bdc-84e9-52b926ab9d66 · outbound

This paper cites Optimal goal-reaching reinforcement learning via quasimetric learning.

Efficient Skill Discovery via Regret-Aware Optimization Optimal goal-reaching reinforcement learning via quasimetric learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.608881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.204893Z digest=sha256:d050871274f297d50d6f7d72fc15928b8cbaf354acd508bbe39cf77d638f502f

Observation 366bb4e3-00e8-4f0d-9c7f-15d14fc0b0d7 · outbound

This paper cites Ski LD : Unsupervised skill discovery guided by factor interactions.

Efficient Skill Discovery via Regret-Aware Optimization Ski LD : Unsupervised skill discovery guided by factor interactions

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.594975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.209112Z digest=sha256:c1d21b0050b735ebe8716fa5da87019ee3e0f150a30d7f2a6ed0b7ebd02b53df

Observation 14c24df4-3b5b-43a6-8166-d494a170eb41 · outbound

This paper cites A comprehensive survey of forgetting in deep learning beyond continual learning.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey of forgetting in deep learning beyond continual learning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.581685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.213181Z digest=sha256:6e52771d377a27609cd2dbc7d19687d54211e47ba20cab3a6d8db3b4af303452

Observation 2d352df8-1244-410b-94cb-5d6481ba4334 · outbound

This paper cites Neural Program Synthesis By Self-Learning.

Efficient Skill Discovery via Regret-Aware Optimization Neural Program Synthesis By Self-Learning

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.271036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.217400Z digest=sha256:f795e7f6ded7c124c8d31334f1c3bbbb682ce6b0d1b23726177d4d46479d8695

Observation 499e79d3-9c9e-4f05-ba46-034c0a95d1b5 · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Behavior contrastive learning for unsupervised skill discovery

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.568005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.222361Z digest=sha256:94a44a82bdbef8ba178ebb609400bb16c3eeba094a5402d797e0e1cfc4ef5671

Observation 9db995b6-e3eb-4e5b-9d9b-30ce1be6506b · outbound

This paper cites Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.553086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.226237Z digest=sha256:ce55b62a0c2da63801241ce9483ef4b01d993e6529748772359ece41f7f17b07

Observation 6c028a05-0532-41b4-8993-8abd49123dec · outbound

This paper cites Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach.

Efficient Skill Discovery via Regret-Aware Optimization Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.538713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.230312Z digest=sha256:2a30690ce88ea19a014575597aa3c1fd29a0e6ad6d769a0ed2f1c3cf60f7632a

Pith citing papers

No inbound Pith citation observations are available.