Pith. sign in

Paper Citation Record · LEDGER

Efficient Skill Discovery via Regret-Aware Optimization

As of 7 August 2026, this Paper Citation Record lists 74 of 74 outbound references and 0 inbound Pith citation observations for arXiv:2506.21044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21044 v1

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:41:38.230312Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

74 of 74 outbound references displayed

  • verified exact2
  • verified fuzzy51
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 90c85963-0971-41b0-9aa5-55b627618389 · outbound

This paper cites write newline.

Efficient Skill Discovery via Regret-Aware Optimization write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.145979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.145979Z digest=sha256:e6e38780372ea8af72c56c2ce712ae2e6a1a46e22d26de90447de6044d29fde1

Observation 1da8fa44-ed1a-4839-883e-e382ad82a665 · outbound

This paper cites M., Crump, T., and Far, B.

Efficient Skill Discovery via Regret-Aware Optimization M., Crump, T., and Far, B

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.322974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.201982Z digest=sha256:ceb1df7813cd650192544f83960746ad53d34273dddad10aee2774e45f7266ab

Observation b46a063f-888c-427f-b844-28504c636e07 · outbound

This paper cites Hindsight Experience Replay.

Efficient Skill Discovery via Regret-Aware Optimization Hindsight Experience Replay

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.306850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.306850Z digest=sha256:92a998a6fd84efa9ec496cfc1d284b18f49a929940b201cc0723d44abb7de9a5

Observation 88f1255f-08d8-45f4-a571-a72faee19108 · outbound

This paper cites TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations.

Efficient Skill Discovery via Regret-Aware Optimization TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:33.424348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:33.424348Z digest=sha256:9814a828dfba0e9096399576aec1202afc248bcbba2e44b7d3f499a016f24507

Observation 1ea52e5c-7a75-40e5-8676-473c33a22ed3 · outbound

This paper cites K., and Konidaris, G.

Efficient Skill Discovery via Regret-Aware Optimization K., and Konidaris, G

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.310533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.518360Z digest=sha256:d74fd5aad7687491c43a695b2a4a64a7f25557ab9221043455aba28dd7df926f

Observation 0e414a57-7eaf-4db7-9a17-26479fb26d17 · outbound

This paper cites Constrained ensemble exploration for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Constrained ensemble exploration for unsupervised skill discovery

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.297553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.609290Z digest=sha256:1f8357b196b0d186fcb9cbfb0343474d05c9561f2be46b96b60f166206467e89

Observation 9a4094e2-9f68-4cd4-bdfe-ea633f5bcc5d · outbound

This paper cites X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U.

Efficient Skill Discovery via Regret-Aware Optimization X., Tanner, J., Vuong, Q., Walling, A., Wang, H., and Zhilinsky, U

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.283764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.703614Z digest=sha256:21a39235f7cf5f395612c8751eec15745926f43867846a45a92da027ec6407c6

Observation 0f826a26-1c03-4767-ade4-ceeb660bfa10 · outbound

This paper cites Exploration by random network distillation.

Efficient Skill Discovery via Regret-Aware Optimization Exploration by random network distillation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.270252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.785283Z digest=sha256:d2c493f3d01e425c14706a04ab2f049b7b4a8e2dde28f04cfa7a927c63a692c0

Observation 55067e30-d11b-4a7b-9452-07042e042327 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills.

Efficient Skill Discovery via Regret-Aware Optimization Explore, discover and learn: Unsupervised discovery of state-covering skills

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.256300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.836290Z digest=sha256:06c5588f8464f533d1f55ffb8cba0100d794079d9f7e0e6f6facddef4640c4c2

Observation 6f836df7-b49c-4975-87b2-7d70de188e82 · outbound

This paper cites Language as a cognitive tool to imagine goals in curiosity driven exploration.

Efficient Skill Discovery via Regret-Aware Optimization Language as a cognitive tool to imagine goals in curiosity driven exploration

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.241660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:33.917944Z digest=sha256:4facf0448df219c1ebac55ede6552102fadaa109e8aca90c613955723612f833

Observation 0b03d062-5f28-41ad-bee5-78a4176256fd · outbound

This paper cites Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey.

Efficient Skill Discovery via Regret-Aware Optimization Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.227451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.005894Z digest=sha256:63890b42a84e47c30ee8243d1aa740351bfa28c8fc11168e8a73cfaaa4fb3e59

Observation ba3fe812-2a9d-4457-abdc-473cae2bb186 · outbound

This paper cites Goal-conditioned imitation learning.

Efficient Skill Discovery via Regret-Aware Optimization Goal-conditioned imitation learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.212302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.090668Z digest=sha256:1aca9d7bfdfab962b6be81b3db39a1129e5b604ce284e0eee7e1bb2a2641893f

Observation a5494acd-cb81-4cd8-a6e5-f246fab0f8bd · outbound

This paper cites Adversarial intrinsic motivation for reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Adversarial intrinsic motivation for reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.198143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.203128Z digest=sha256:b70c8aef08ea35e4ecd9a40b3ef31ef2221b1035bbe0998281a5e43c766dcfe8

Observation e284d72f-c053-4426-b2b3-0abe976a1ce9 · outbound

This paper cites Diversity is all you need: Learning skills without a reward function.

Efficient Skill Discovery via Regret-Aware Optimization Diversity is all you need: Learning skills without a reward function

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.184132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.247740Z digest=sha256:364ab447d061f01a78bbf5085a7b46e91d40f2a520a4f481e25e861777c0410e

Observation 143187d5-3f4d-4258-9ef2-695f91139aa3 · outbound

This paper cites C-Learning: Learning to Achieve Goals via Recursive Classification.

Efficient Skill Discovery via Regret-Aware Optimization C-Learning: Learning to Achieve Goals via Recursive Classification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.332286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.332286Z digest=sha256:743f80b03d8207eaf868fc66cc96dc7abb1b71fd5875ec2ac16e85fc72011071

Observation 20c0ea5b-129c-4409-bda9-7de5ec4a561b · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.170092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.417950Z digest=sha256:76b2eb71ba034fe53e37c1f4028a79dc63c5e0b5fcf728d280134f2ab23c280a

Observation d9147cd4-2c24-4fa6-b982-d8eec0274326 · outbound

This paper cites Curriculum-guided hindsight experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Curriculum-guided hindsight experience replay

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.156692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.468891Z digest=sha256:2fc4b97f0adea1d1acb4b29f5ed1a685fed1939f60f1a141e99e556b077e16d0

Observation a5e18089-2bb6-4ebf-b2dd-aa137296a80a · outbound

This paper cites Automatic goal generation for reinforcement learning agents.

Efficient Skill Discovery via Regret-Aware Optimization Automatic goal generation for reinforcement learning agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.473619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.473619Z digest=sha256:959cad7097aafced9c888fd470826feceb59608a6045a2b6ba39e9f2436c11fd

Observation 7d6cf578-631c-4d71-849a-32e0cf78ae83 · outbound

This paper cites Accuracy-based Curriculum Learning in Deep Reinforcement Learning.

Efficient Skill Discovery via Regret-Aware Optimization Accuracy-based Curriculum Learning in Deep Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.578090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.578090Z digest=sha256:e5ba6a28d0f4026067aae09c3b7073f251183a83d544f470bb69acce9aaba6e1

Observation ccb63044-06ba-4f5f-8633-324719177297 · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2020.

Efficient Skill Discovery via Regret-Aware Optimization D4rl: Datasets for deep data-driven reinforcement learning, 2020

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.133429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:34.762057Z digest=sha256:a9d711f1a356860faecc8432b715d0208e4a143ba6f862aac90c2ae0c269b304

Observation 77335ca6-e087-4292-8657-b2892f585840 · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Efficient Skill Discovery via Regret-Aware Optimization Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.963486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.963486Z digest=sha256:973b67006e9b9c4a5165ef58b872fa80bcdec190522153fcca2d7d6f5b2efb8d

Observation 53a1012f-57ab-4f05-a980-b04548088400 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.121608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.121608Z digest=sha256:8ea5fe54d9aa451a6d25aa5732b2fc97f75aabdad347d76a83fef73b9a653768

Observation 50abf721-f062-4549-8f28-ca8c2852b4eb · outbound

This paper cites J., and Wierstra, D.

Efficient Skill Discovery via Regret-Aware Optimization J., and Wierstra, D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.110596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.261687Z digest=sha256:2c22682dd78d6713e50f783f0a109eed9895c59cc2c76ab4f4bddbd00e4109df

Observation 37668b98-4ab1-4136-be9f-db203f06ca2b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Efficient Skill Discovery via Regret-Aware Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:35.471301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:35.471301Z digest=sha256:564b0d04933572bb71b05ba1ab96b142d6f96e6e19f6dfe5026d020d515760ce

Observation a1cbfc2f-458d-472e-a9bb-20c95be91d63 · outbound

This paper cites Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation.

Efficient Skill Discovery via Regret-Aware Optimization Emergence of Structured Behaviors from Curiosity-Based Intrinsic Motivation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.461935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.692807Z digest=sha256:8e2b5a8f0840bf40cc17c3d7bc09c8c05a1a189e0a514fbc4fe60e9678668b53

Observation 33de2148-99d3-48aa-8870-312b786ce215 · outbound

This paper cites Exploration in deep reinforcement learning: From single-agent to multiagent domain.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: From single-agent to multiagent domain

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.089016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:35.858079Z digest=sha256:6df16ae7efa8819f9d6fbee52989166d935354b4e64add9f8699ba04991b139b

Observation 3a774f42-b30f-4306-8538-3989effc3200 · outbound

This paper cites Open-Endedness is Essential for Artificial Superhuman Intelligence.

Efficient Skill Discovery via Regret-Aware Optimization Open-Endedness is Essential for Artificial Superhuman Intelligence

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.029860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.029860Z digest=sha256:43a9cb26c82533b4653f2f50795254df434680d4f89fa97668a99b158e7f418f

Observation ce161550-23c4-49d2-a496-0089d72ad128 · outbound

This paper cites Unsupervised curricula for visual meta-reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised curricula for visual meta-reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.074905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.228723Z digest=sha256:e528f05087e2d10f1256ac85205d9bf37ef6bfd1def804f133a59f1352493d7b

Observation 63d64998-cf2b-468f-9fe6-c4c89b7cb46e · outbound

This paper cites A comprehensive survey on self-interpretable neural networks.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey on self-interpretable neural networks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.398189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.398189Z digest=sha256:104ce00be80a93a53b0b3edeac102de67dc2aa310dd5d3e8c8c2708a91ec2869

Observation f843fe81-8fb2-4a69-a52a-25fa443fe92a · outbound

This paper cites Replay-Guided Adversarial Environment Design.

Efficient Skill Discovery via Regret-Aware Optimization Replay-Guided Adversarial Environment Design

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:41:38.328725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.601070Z digest=sha256:1fe514aebe64eb5892e5f9028366951a30a0ed8e2ea4fe5138856412b619f4ae

Observation 1fabd924-f58b-4583-b50b-12a9223ad4f0 · outbound

This paper cites Prioritized level replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized level replay

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.061604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.669110Z digest=sha256:cdb49f3fbf26d911de8a03703b1d7dc771bc68fea46e1f366f94ef52ae107a53

Observation d4f15ff1-9b76-4682-98e1-2f2825d64cbe · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:39.047719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.714340Z digest=sha256:25585f78b607454dd146b70129d331e7d8571a3ebbf1bffae6e51100d955946a

Observation 844b097e-fc4a-44d9-959f-3033e3aa5941 · outbound

This paper cites and Oudeyer, P.-Y.

Efficient Skill Discovery via Regret-Aware Optimization and Oudeyer, P.-Y

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:36.740268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:36.740268Z digest=sha256:d81be39a1aae1550f79f241e02ffa2931d84ca572d82fd489f7b1bf8e8c0926f

Observation 3b147d4b-9840-417d-a8ce-3fcd0f63d7b5 · outbound

This paper cites K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J.

Efficient Skill Discovery via Regret-Aware Optimization K., Lee, H., Hwang, D., Park, S., Min, K., and Choo, J

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.026333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.814734Z digest=sha256:9335a7735dd267656038d9f11d8017960132439e9cf641b67d135d4282dc9fa5

Observation e96c04cd-b09c-4589-87e6-e8d3b65bec21 · outbound

This paper cites Unsupervised skill discovery with bottleneck option learning, 2021.

Efficient Skill Discovery via Regret-Aware Optimization Unsupervised skill discovery with bottleneck option learning, 2021

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:39.012568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.865430Z digest=sha256:f24cde23a1fa0866af2ba5c4adc99a527d8f8d73beb616fb494b37675290c882

Observation 31c84aeb-d9aa-468b-842f-3722166e7392 · outbound

This paper cites Active world model learning with progress curiosity.

Efficient Skill Discovery via Regret-Aware Optimization Active world model learning with progress curiosity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.998539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:36.942087Z digest=sha256:f744c9463629ec4c7437ac95fb3ab37f2fd57f94acf584abd62e96868714654b

Observation 661f3f1b-5f3b-4fb8-947c-aef0e73963a1 · outbound

This paper cites Exploration in deep reinforcement learning: A survey.

Efficient Skill Discovery via Regret-Aware Optimization Exploration in deep reinforcement learning: A survey

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.985466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.015977Z digest=sha256:6926a79c038d8bfb317e490f4191bea5cbaab1c308d3a2c3b839b606ae2dc8fb

Observation b6d116a6-a48f-4cab-8144-738129961de2 · outbound

This paper cites B., Yarats, D., Rajeswaran, A., and Abbeel, P.

Efficient Skill Discovery via Regret-Aware Optimization B., Yarats, D., Rajeswaran, A., and Abbeel, P

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.971539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.079101Z digest=sha256:725eeb36824107d6d8ce1bc44ee272b39cf51ad261cf80de33f6d028240edfab

Observation 7fb3eaea-27e2-4d7f-b4c3-a3b8e91864d3 · outbound

This paper cites and Seo, S.-W.

Efficient Skill Discovery via Regret-Aware Optimization and Seo, S.-W

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.957700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.143371Z digest=sha256:7c6cf41eabc06eb640136614466f43b66b1c0443e4c76573a6e0a17e5f3bba05

Observation 95fd50be-07ad-4499-b263-8839f0ba694b · outbound

This paper cites Revisiting graph adversarial attack and defense from a data distribution perspective.

Efficient Skill Discovery via Regret-Aware Optimization Revisiting graph adversarial attack and defense from a data distribution perspective

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.944068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.200699Z digest=sha256:00490708ba366637782d62b5d5f074bb6d62581159852bd9bcd416e8435d8e7e

Observation 5cbd798b-7508-4414-98dd-1f02158d4e1f · outbound

This paper cites Boosting the adversarial robustness of graph neural networks: An ood perspective.

Efficient Skill Discovery via Regret-Aware Optimization Boosting the adversarial robustness of graph neural networks: An ood perspective

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.929948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.273877Z digest=sha256:256cee3ba9fa2a957621f92892b54b258319d4622277dc122c41daba39dd7fe7

Observation 8972b79e-4f56-4c4d-879d-e3ff63f6d453 · outbound

This paper cites A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals.

Efficient Skill Discovery via Regret-Aware Optimization A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:37.321403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:37.321403Z digest=sha256:468484ca4c7d0c5c9738a9109feaa276719ad2f0e0c68d018b18a65286058368

Observation 17c92cf3-8124-4217-8929-958f476c3f8b · outbound

This paper cites Choreographer: Learning and adapting skills in imagination.

Efficient Skill Discovery via Regret-Aware Optimization Choreographer: Learning and adapting skills in imagination

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.916681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.365884Z digest=sha256:b37edddd80ca5666acb0ea78fd00bad71bebdbb2f3d8a837086f31d337aff567

Observation 57d2ebd6-1b94-4784-b061-afed5604589b · outbound

This paper cites Planning with goal-conditioned policies.

Efficient Skill Discovery via Regret-Aware Optimization Planning with goal-conditioned policies

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.901968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.462003Z digest=sha256:62924c53bd1f30c0068ac72a138430dcd5bb739e41b30a3f2fc89d0b3edb1b5f

Observation 06538529-6818-4526-94fe-164caa8d0a93 · outbound

This paper cites Wasserstein dependency measure for representation learning.

Efficient Skill Discovery via Regret-Aware Optimization Wasserstein dependency measure for representation learning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.889465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.499235Z digest=sha256:376840c21582da39b0cf0ccf39089f400a42da6cd2de1b0dfef149dc314c3c4c

Observation cb1c474e-9636-4702-886b-d7f66500ee3f · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Lipschitz-constrained unsupervised skill discovery

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.876381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.582284Z digest=sha256:42f940188d31c287186590217ebcfffa58d5c8ed73adb58041c602cf3229d0c8

Observation 2318ca52-4a84-4898-b339-788fcc50afb8 · outbound

This paper cites Hiql: Offline goal-conditioned rl with latent states as actions.

Efficient Skill Discovery via Regret-Aware Optimization Hiql: Offline goal-conditioned rl with latent states as actions

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.863376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.652587Z digest=sha256:f5ef4aea2ded5ffdb5ae5db613858da6c81eb15701e37a6b93d34030ffbcb6bb

Observation edf572e2-741e-428e-8fa2-e7cf5223703b · outbound

This paper cites Foundation policies with hilbert representations.

Efficient Skill Discovery via Regret-Aware Optimization Foundation policies with hilbert representations

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.849816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.697677Z digest=sha256:1d2418eb6527a63d838799c57973a14e869f239034b427457cc61453a38084bd

Observation 1c666100-dff7-49e6-9bb4-4666000b500c · outbound

This paper cites METRA : Scalable unsupervised RL with metric-aware abstraction.

Efficient Skill Discovery via Regret-Aware Optimization METRA : Scalable unsupervised RL with metric-aware abstraction

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.837059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.797162Z digest=sha256:bfa19ee65e078b0b73b665b339a34238f534ab8246dc70a7ea0589c3adcbe633

Observation b04ca4b0-c24e-47a8-a6e3-beb4780880fa · outbound

This paper cites Evolving curricula with regret-based environment design.

Efficient Skill Discovery via Regret-Aware Optimization Evolving curricula with regret-based environment design

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.824248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.867902Z digest=sha256:55c313ec9df07e4a80fb39edd11d1afc531e53ba74c0f70dafc092ce40bd60db

Observation 06fc72e3-eb0c-41b2-aeec-450f54d74c22 · outbound

This paper cites A., and Darrell, T.

Efficient Skill Discovery via Regret-Aware Optimization A., and Darrell, T

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.810397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:37.943248Z digest=sha256:7622ac777eb70dc207cbde01a030c789bd5f0f76e69933694664543c4e3b511d

Observation 99fba55d-032c-4e6b-9bbb-6e0f2bb19b85 · outbound

This paper cites H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S.

Efficient Skill Discovery via Regret-Aware Optimization H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.796256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.036854Z digest=sha256:e0f6f31d923e9c7e1ad2f0dcdfa9581e7c5e163155259416a58310f132c0f3e5

Observation 9cfa1ce2-4d30-42d0-8d0b-d85150d118b1 · outbound

This paper cites Automatic Curriculum Learning For Deep RL: A Short Survey.

Efficient Skill Discovery via Regret-Aware Optimization Automatic Curriculum Learning For Deep RL: A Short Survey

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.079354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.079354Z digest=sha256:14e6310b73fa1f4707d75d4dc0590d7aa4c8c4d79ad46215560d1d28fcdd02fe

Observation 428819a6-0947-4e3d-90bb-107c5f699bdb · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:41:38.783134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.121688Z digest=sha256:b80136380f94756a9d45128c911e88699dde85dfc494784ca4a00a2c05a97c0c

Observation 20125f29-6832-4238-89d4-625fbd6c8abc · outbound

This paper cites Prioritized experience replay.

Efficient Skill Discovery via Regret-Aware Optimization Prioritized experience replay

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.770421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.131699Z digest=sha256:69a90e7b065c8eda677c0257acd4dd5cda47e0832c6e78b07020a15e14a3156b

Observation c49c44bb-ec1b-4384-93b5-5b9283d4a7b8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Proximal Policy Optimization Algorithms

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.139133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.139133Z digest=sha256:b5ee11dff67f3b30d182a123d770b382d293e0a11b2456aec9f779d96deb24c4

Observation 37147148-54ea-48b5-ae05-9d0211ef2774 · outbound

This paper cites Dynamics-aware unsupervised discovery of skills.

Efficient Skill Discovery via Regret-Aware Optimization Dynamics-aware unsupervised discovery of skills

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.755575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.149622Z digest=sha256:ddec72175ead1d288f3b02465fccff116c6b95c5f0d38debc7528b58507342ba

Observation 3ece75c2-8146-4a87-aac1-db7807dc5998 · outbound

This paper cites Deterministic policy gradient algorithms.

Efficient Skill Discovery via Regret-Aware Optimization Deterministic policy gradient algorithms

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.740517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.157683Z digest=sha256:c06567e3446ce9d0ac8c47e97e1d75ee80462c0d2e9e80c1278a41462f444efc

Observation ae196208-c010-45c4-9dee-eea8293b97e0 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play.

Efficient Skill Discovery via Regret-Aware Optimization A general reinforcement learning algorithm that masters chess, shogi, and go through self-play

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.725250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.165986Z digest=sha256:9a53c7a64cf2b7b50525002a37499874b4d0e0db6ddeeb91523c9cd1cadb8dcf

Observation 05395d97-732b-444f-971b-fd92e2885772 · outbound

This paper cites Intrinsic motivation and automatic curricula via asymmetric self-play.

Efficient Skill Discovery via Regret-Aware Optimization Intrinsic motivation and automatic curricula via asymmetric self-play

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.710934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.170823Z digest=sha256:1074f3dafecb002caff4cbdd359f8775a4e732f0c66e3f8165b6e5253aec54dc

Observation 72c02ece-e109-43a7-a90a-62125a6fec8f · outbound

This paper cites Policy continuation with hindsight inverse dynamics.

Efficient Skill Discovery via Regret-Aware Optimization Policy continuation with hindsight inverse dynamics

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.696597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.174661Z digest=sha256:2ecb72a8fc4521222b5839013554cbcffc9812bd989ed8f5bbfe458416d8350b

Observation 9343b173-20df-48ed-9c25-1f9d65641445 · outbound

This paper cites Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections.

Efficient Skill Discovery via Regret-Aware Optimization Hierarchical reinforcement learning for dynamic autonomous vehicle navigation at intelligent intersections

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.682455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.179067Z digest=sha256:e5c17a41c3f501dfd9fd1aaa3ec9d01833faee5a59ca4cb62f6d8d2e16f46ea1

Observation fdad43a7-eb70-470d-baff-90cdd4c02170 · outbound

This paper cites Market-aware long-term job skill recommendation with explainable deep reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Market-aware long-term job skill recommendation with explainable deep reinforcement learning

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.667895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.183587Z digest=sha256:eef33e47fc349515d1671760ebc45df3f1089157d6e84fc6870fa22e336e6c3d

Observation 8abb6957-0f84-4e6a-b131-b2d2787a6ba8 · outbound

This paper cites an unresolved cited work.

Efficient Skill Discovery via Regret-Aware Optimization Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.188231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.188231Z digest=sha256:5e617f04dc7bde906d365703a1a8a54aa4aa68b448fe459ace808602c53849e3

Observation 2c1ae144-c4ff-4eb7-bd6d-223d2bb13244 · outbound

This paper cites S., McAllester, D., Singh, S., and Mansour, Y.

Efficient Skill Discovery via Regret-Aware Optimization S., McAllester, D., Singh, S., and Mansour, Y

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:38.191955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:38.191955Z digest=sha256:130e8b0e36d3ffbcc2ac8137c3c5ff7025d6c78ff337469b34088e474ce465d9

Observation 0c85fde9-bacb-4a1b-8a53-ba00562a2777 · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Efficient Skill Discovery via Regret-Aware Optimization dm\_control: Software and tasks for continuous control

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.637023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.196846Z digest=sha256:774906ade49f3ad78ce464fbdbfa9437c51a76e4d28fb55d72fb8cf51ab895ab

Observation 7fe00edc-7e18-485c-8903-9a3394f52bc8 · outbound

This paper cites M., Mathieu, M., Dudzik, A., Chung, J., Choi, D.

Efficient Skill Discovery via Regret-Aware Optimization M., Mathieu, M., Dudzik, A., Chung, J., Choi, D

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.623333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.201024Z digest=sha256:3a3f5478b19aba10fb81df9679c42f5d62c28d0a5f889499058df02606518e08

Observation d7b4d0c5-2d5f-4bdc-84e9-52b926ab9d66 · outbound

This paper cites Optimal goal-reaching reinforcement learning via quasimetric learning.

Efficient Skill Discovery via Regret-Aware Optimization Optimal goal-reaching reinforcement learning via quasimetric learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.608881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.204893Z digest=sha256:16ced7e49d77b389c79989397b23f585b71a6fe8d1fa960908893a20c5d79347

Observation 366bb4e3-00e8-4f0d-9c7f-15d14fc0b0d7 · outbound

This paper cites Ski LD : Unsupervised skill discovery guided by factor interactions.

Efficient Skill Discovery via Regret-Aware Optimization Ski LD : Unsupervised skill discovery guided by factor interactions

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.594975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.209112Z digest=sha256:2d3033c4385ad7bcd9861ac0f06184412da8a8dd4a088a1f3e9ab5391be09094

Observation 14c24df4-3b5b-43a6-8166-d494a170eb41 · outbound

This paper cites A comprehensive survey of forgetting in deep learning beyond continual learning.

Efficient Skill Discovery via Regret-Aware Optimization A comprehensive survey of forgetting in deep learning beyond continual learning

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.581685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.213181Z digest=sha256:4e30679fe06b2c6a24ae9f29bc5510d9d855df6b7dc7cd61973c7c6f088cfffa

Observation 2d352df8-1244-410b-94cb-5d6481ba4334 · outbound

This paper cites Neural Program Synthesis By Self-Learning.

Efficient Skill Discovery via Regret-Aware Optimization Neural Program Synthesis By Self-Learning

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:41:38.271036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.217400Z digest=sha256:0722ccc9ac15b208071cd94eda0960aee9630d604887ddadcc4430421c1414e6

Observation 499e79d3-9c9e-4f05-ba46-034c0a95d1b5 · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery.

Efficient Skill Discovery via Regret-Aware Optimization Behavior contrastive learning for unsupervised skill discovery

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.568005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.222361Z digest=sha256:6bd242555a50eb85cf9f19cc18e5df344543e939b25807effa1e8411089f5ebb

Observation 9db995b6-e3eb-4e5b-9d9b-30ce1be6506b · outbound

This paper cites Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning.

Efficient Skill Discovery via Regret-Aware Optimization Interactive interior design recommendation via coarse-to-fine multimodal reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.553086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.226237Z digest=sha256:9a5961e7e9699dcd2ee156d3281b06714ee604fd075cc3aae205c608beba6967

Observation 6c028a05-0532-41b4-8993-8abd49123dec · outbound

This paper cites Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach.

Efficient Skill Discovery via Regret-Aware Optimization Generative learning plan recommendation for employees: A performance-aware reinforcement learning approach

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:41:38.538713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T22:41:38.230312Z digest=sha256:9096b31bfbdb5c136de153ea8dd0b5568fe2a9f0dee140925925a1ca4bdaa75a

Pith citing papers

No inbound Pith citation observations are available.