Pith. sign in

Paper Citation Record · LEDGER

Unsupervised Skill Discovery through Skill Regions Differentiation

As of 8 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2506.14420.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14420 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:27:00.722538Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 71f17265-6779-4c69-a3c3-c18f62a0d420 · outbound

This paper cites A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,.

Unsupervised Skill Discovery through Skill Regions Differentiation A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.400282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.400282Z digest=sha256:a4d3641c7f30ca6a31fc044d0665e70aae518c403446650c0195cb95a0c9262d

Observation ef1d3c3a-ae07-49e5-adc8-6e5fb3e25493 · outbound

This paper cites Mastering atari games with limited data,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering atari games with limited data,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.370724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.406553Z digest=sha256:1c3a8cac258450b1dc1200e1c360b221bb37a9dcdcc71d1b8b72ad20cb25bc8e

Observation b2ec47ef-7373-48d0-a232-25be158a95b7 · outbound

This paper cites Continu- ous improvement of self-driving cars using dynamic confidence-aware reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Continu- ous improvement of self-driving cars using dynamic confidence-aware reinforcement learning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.360799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.410591Z digest=sha256:47eb406a904cd111eae776b78430cc1bf15a4715f1178e6af1c3221a82429061

Observation 8394700b-fd26-46cf-b2d3-dd2159d76866 · outbound

This paper cites Uncertainty-aware model-based reinforce- ment learning: Methodology and application in autonomous driving,.

Unsupervised Skill Discovery through Skill Regions Differentiation Uncertainty-aware model-based reinforce- ment learning: Methodology and application in autonomous driving,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.350983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.415890Z digest=sha256:049cf7e44d12cdf5a430a65f5a076773d34bc07bdc3062c4da14d054f4b22c68

Observation 704b35af-edb4-4fd9-9afc-82911f0acd5c · outbound

This paper cites Temporal difference learning for model predictive control,.

Unsupervised Skill Discovery through Skill Regions Differentiation Temporal difference learning for model predictive control,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.340781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.421942Z digest=sha256:7c0e7df54800792f817f32ccd7a70fe70a6ff8109c45f2b6467315859436f494

Observation ed24a445-4e3a-4971-a598-29b233b89200 · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning robust perceptive locomotion for quadrupedal robots in the wild,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.330608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.426007Z digest=sha256:fdf37bde833bf709b624bcce4c113f6c49d8794e9d691d7b44371014d100c141

Observation 8dde356a-dbf6-43df-8776-353702a37f9c · outbound

This paper cites Relay hindsight experience replay: Self-guided continual reinforcement learning for sequential object manipulation tasks with sparse rewards,.

Unsupervised Skill Discovery through Skill Regions Differentiation Relay hindsight experience replay: Self-guided continual reinforcement learning for sequential object manipulation tasks with sparse rewards,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.430953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.430953Z digest=sha256:574533ea7bf91531f54641fd10049e7366dfe5b4b78c62759e3294c6de263861

Observation e58e16ff-5cdf-4271-b646-8e358a74f9e9 · outbound

This paper cites Reward design with language models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reward design with language models,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.314030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.434350Z digest=sha256:3c92014c31948e76d55e3938981968bef822aef7b8d9e35bbb56b2455fbb401a

Observation d4c1edc9-f530-46c2-bf7b-c3e216841b4f · outbound

This paper cites Maniskill2: A unified benchmark for generalizable manipulation skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Maniskill2: A unified benchmark for generalizable manipulation skills,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.303638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.438282Z digest=sha256:2e3c8618350d09578d06f6056067a723b286b7b3ad9ab6bddf8ebe506c89b07c

Observation 619f9568-8743-48ab-b70a-aa94a8152069 · outbound

This paper cites Pre-trained models: Past, present and future,.

Unsupervised Skill Discovery through Skill Regions Differentiation Pre-trained models: Past, present and future,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.293541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.443129Z digest=sha256:76767cebe0cf63457907148d0c7c3e32be181a2784e2c40709d95f079539037e

Observation 4fd48582-3036-4fbe-860a-e6ca0ba79d71 · outbound

This paper cites GPT-4 Technical Report.

Unsupervised Skill Discovery through Skill Regions Differentiation GPT-4 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.448023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.448023Z digest=sha256:aeba852fb8f9aec74121eac432da0d122ae7b619ed7885e03caa75f276722d24

Observation a57c93fb-c2b9-415d-9369-dde7bd856de2 · outbound

This paper cites Training language models to follow instructions with human feedback,.

Unsupervised Skill Discovery through Skill Regions Differentiation Training language models to follow instructions with human feedback,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.452438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.452438Z digest=sha256:106801626d8bcf90e45bb136ba4e327c780245c8187cc6a70f39211ff366f7c6

Observation 6011ef50-a54c-4a95-9f8a-97e8b97c3e0f · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Unsupervised Skill Discovery through Skill Regions Differentiation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.456970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.456970Z digest=sha256:3dc94963ff47f26b7ed1761e57e1dc6be9b80ca1c0576bef8137328c7af6ca9f

Observation 226fe321-0ecf-4b4d-aff0-bbe46e7b5c6e · outbound

This paper cites Masked au- toencoders are scalable vision learners,.

Unsupervised Skill Discovery through Skill Regions Differentiation Masked au- toencoders are scalable vision learners,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.461759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.461759Z digest=sha256:3eab9afa2ff6b77e82900885d1ddbd3e1f0a9934c9fe38339e8cafa37a1ff39b

Observation fe02bf74-1ab5-4135-b72a-c480516daa08 · outbound

This paper cites V-JEPA: Latent video prediction for visual representation learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation V-JEPA: Latent video prediction for visual representation learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.268456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.465949Z digest=sha256:91220670e5fbdc50dfe389615c437845c3639e921d34264f602d372941b9ad10

Observation b20b188b-cf5e-4a2c-8a52-136036aad63f · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Bootstrap your own latent-a new approach to self-supervised learning,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.472353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.472353Z digest=sha256:579b254380163d420530715ec6783fae06ceb677340981c1d4a9a22e7d463aa6

Observation 30b3293a-b4c1-4a8c-83b1-2ed348eb81f4 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

Unsupervised Skill Discovery through Skill Regions Differentiation Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.250413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.478396Z digest=sha256:e9c547c7570032639802d5899d13411d77addc8d0464c1d3a8bd50230d0de0a3

Observation 220c38ac-5d7d-4b90-bcd2-44cddea60b37 · outbound

This paper cites R3m: A universal visual representation for robot manipulation,.

Unsupervised Skill Discovery through Skill Regions Differentiation R3m: A universal visual representation for robot manipulation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.240433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.485375Z digest=sha256:8994d0818893897b9e80981df90ea93b995904c64f4006d9d57758caf70ce9e3

Observation 36d4a9dc-19eb-4d33-801a-cf0dfc0a383f · outbound

This paper cites URLB: Unsupervised reinforcement learning benchmark,.

Unsupervised Skill Discovery through Skill Regions Differentiation URLB: Unsupervised reinforcement learning benchmark,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.230131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.489329Z digest=sha256:40da21da85afb9ac2b790b9690f634fbb17811de53f8b1c4b23d835c4d1df0a3

Observation 904361d9-e69c-40ef-bd35-575853178d1d · outbound

This paper cites Variational Intrinsic Control.

Unsupervised Skill Discovery through Skill Regions Differentiation Variational Intrinsic Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.497698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.497698Z digest=sha256:fd9ed9ddf1785b544f0dc48b136d7f5180e86fd5078e8d0e21f808f2e3d0bfa4

Observation 7b6e1eed-5803-4e2e-8b8c-5842aef46ac9 · outbound

This paper cites Behavior from the void: Unsupervised active pre-training,.

Unsupervised Skill Discovery through Skill Regions Differentiation Behavior from the void: Unsupervised active pre-training,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.219172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.501331Z digest=sha256:3965014b5c1a6a4e9c72591725297049b94121adb8dd142188bf9d9cca2f736e

Observation a619baad-5a87-4765-9ed0-96b73d1ae417 · outbound

This paper cites Understanding the limitations of variational mutual information estimators,.

Unsupervised Skill Discovery through Skill Regions Differentiation Understanding the limitations of variational mutual information estimators,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.207390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.505838Z digest=sha256:4d84ea988504a1a288d4be1bea7d0f11b7e077e4053cd1f58d8cd082b3ab314c

Observation 310b62bc-be25-4cdf-96d8-474153cd3b5c · outbound

This paper cites Diversity is all you need: Learning skills without a reward function,.

Unsupervised Skill Discovery through Skill Regions Differentiation Diversity is all you need: Learning skills without a reward function,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.196242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.510281Z digest=sha256:bf396fa20265a05bf57e49adaff9e467e2470f0cbd4bfbba971455bb22e60a51

Observation 0eaa15a3-04d2-4e2a-b9e0-2a1cf9db9cdc · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Behavior contrastive learning for unsupervised skill discovery,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.185751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.514388Z digest=sha256:a02c9388edfd420f652724f124acd68a9f87d40853bf68fb534b8e6f1fe14ef0

Observation 5f7cea8d-1147-4b16-82f1-665d802b6b56 · outbound

This paper cites METRA: Scalable unsupervised RL with metric-aware abstraction,.

Unsupervised Skill Discovery through Skill Regions Differentiation METRA: Scalable unsupervised RL with metric-aware abstraction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.174666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.520823Z digest=sha256:1d9141c0957f7141aa08cc8f0659a5c13c6a1043e0a2af171f46cae1bb420eec

Observation 455f1e32-4fc6-44e9-a00c-3ae7ae9e7605 · outbound

This paper cites Lipschitz-constrained unsupervised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Lipschitz-constrained unsupervised skill discovery,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.530003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.530003Z digest=sha256:8e5c9b8db95f1bcd10c99271d59b75a76f18b1c167ecba499c43c69ead18bbc7

Observation 8e11a816-fa34-4e9e-9f6b-eb50d1a3912a · outbound

This paper cites Controllability-aware unsuper- vised skill discovery,.

Unsupervised Skill Discovery through Skill Regions Differentiation Controllability-aware unsuper- vised skill discovery,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.157930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.540883Z digest=sha256:291fdc8233dc5d7c9f26e2a6dd3264af2c6ccbcb1474c7b49656395091684c09

Observation 0151c8b4-a341-4e52-add7-ba2c19c31fe1 · outbound

This paper cites Unsupervised reinforcement learning with contrastive intrinsic control,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised reinforcement learning with contrastive intrinsic control,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.146584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.554184Z digest=sha256:255295f4af2e5acdf20da7b9abf69ca181581dcd81be7c19ce6d3502c0e740c4

Observation 9834ffa6-f4e4-4a50-bde2-7cc8dff66686 · outbound

This paper cites Mastering the unsupervised reinforce- ment learning benchmark from pixels,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering the unsupervised reinforce- ment learning benchmark from pixels,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.135541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.565679Z digest=sha256:3c174275c2c6c19095062c3cc8c411026f653d9d3bf2d1c2f88354c0c5218234

Observation eb901812-b8f3-4efb-8a55-47a2b42fc20c · outbound

This paper cites Auto-Encoding Variational Bayes.

Unsupervised Skill Discovery through Skill Regions Differentiation Auto-Encoding Variational Bayes

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.580787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.580787Z digest=sha256:24274e1ca4f3f87dd4acd69c15ca5709c0bc8ef725c93f1141ca2d6b214b8972

Observation 657529dc-3a0d-4c4c-b252-2c7bf0b89cc7 · outbound

This paper cites An introduction to variational autoencoders,.

Unsupervised Skill Discovery through Skill Regions Differentiation An introduction to variational autoencoders,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.594940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.594940Z digest=sha256:26f12eee05e41b3658b1a96b87251a9f18241ae83cdcfb74daa1d60d87e44e39

Observation bf395112-8f86-4dee-8e21-659391494978 · outbound

This paper cites Multi-task reinforcement learning with soft modularization,.

Unsupervised Skill Discovery through Skill Regions Differentiation Multi-task reinforcement learning with soft modularization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.118194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.598765Z digest=sha256:7814aa8c19bdf2e210f3bf450c93fe525afba8acf842f6f3765e256b905ef1ca

Observation 83799067-9158-4cf0-8018-88614f6ce87f · outbound

This paper cites Near-bayesian exploration in polynomial time,.

Unsupervised Skill Discovery through Skill Regions Differentiation Near-bayesian exploration in polynomial time,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.107853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.606260Z digest=sha256:66da5a3f894898da320b4bb03033c99f838e9616f352ac7741c731751ebc4a9e

Observation 772d9560-1c81-45ba-bc53-39168129a0c9 · outbound

This paper cites An analysis of model-based interval estimation for markov decision processes,.

Unsupervised Skill Discovery through Skill Regions Differentiation An analysis of model-based interval estimation for markov decision processes,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.096620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.609774Z digest=sha256:a5be9bd0955336229c5aaead5df30f7953b1095592c9ebbdc293f4161947073e

Observation 2809328a-2103-4dae-befb-55cdcac852fd · outbound

This paper cites Unifying count-based exploration and intrinsic motivation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unifying count-based exploration and intrinsic motivation,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.087603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.612633Z digest=sha256:1e654e84d0adb4a234b9a8eb8fbd2a021c9af3a82aae2ea455fd56a64a94f744

Observation 072ed9ba-5c18-4643-97a3-a90c2c7dcb13 · outbound

This paper cites Count-based exploration with neural density models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Count-based exploration with neural density models,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.078190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.616640Z digest=sha256:323d24f8f9a7dd830aa5c1736e6007f2fc5aaef4a787d9e1c53a048a75360194

Observation 530eb6f2-b55d-4f03-9206-a4563300a6aa · outbound

This paper cites Dynamics- aware unsupervised discovery of skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Dynamics- aware unsupervised discovery of skills,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.068018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.619907Z digest=sha256:27bfc29a8355dd68cb4622e9c31452ac99c96f779def4a39551fe608b8c487af

Observation cfa7a766-a5eb-47db-a842-e327c7ea0c66 · outbound

This paper cites Explore, discover and learn: Unsupervised discovery of state-covering skills,.

Unsupervised Skill Discovery through Skill Regions Differentiation Explore, discover and learn: Unsupervised discovery of state-covering skills,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.058443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.623315Z digest=sha256:91d7c78a523e4ea7cba580c9e0902dafa703b05c11ac3fc37e400c5ade1221ee

Observation 4a3bb0fb-7ccb-49ae-b38b-d94623953078 · outbound

This paper cites Unsupervised skill discovery via recurrent skill training,.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised skill discovery via recurrent skill training,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.048859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.626404Z digest=sha256:650f898993855cc666578c7c26b588706070cb7b9ad371f7cfa4e59fdb5eacb7

Observation 03fd7f24-e87f-48cf-9479-f453ef079ad9 · outbound

This paper cites Learning to discover skills with guidance,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning to discover skills with guidance,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.038583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.629803Z digest=sha256:582fbe043ee68caea5d993cbdd6ebf2957561f38b570bada567447a97e6422ef

Observation 5bc42a72-5855-459d-98d9-059d07f1d69e · outbound

This paper cites Aps: Active pretraining with successor features,.

Unsupervised Skill Discovery through Skill Regions Differentiation Aps: Active pretraining with successor features,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.028714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.633081Z digest=sha256:bc45aae8632f86d41e6f94180de5727fe60a543dac3afce6bd1c23bf2e11e706

Observation da26a86e-3728-495a-9390-35aa84fa09de · outbound

This paper cites Efficient exploration via state marginal matching,.

Unsupervised Skill Discovery through Skill Regions Differentiation Efficient exploration via state marginal matching,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.017875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.636865Z digest=sha256:51df43b13272e84a88fbb771813312d76b1d64fbf9ef6868c223983cf44feea1

Observation f32b9b01-64de-403f-95ca-bded11c4e539 · outbound

This paper cites Learning more skills through optimistic exploration,.

Unsupervised Skill Discovery through Skill Regions Differentiation Learning more skills through optimistic exploration,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.997823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.644518Z digest=sha256:349756dd24f20e4c85faacf835e037be986a737b63d9867a2fe428719c132c21

Observation f6728e22-9bc1-4fde-bf6b-b38aeb7858cf · outbound

This paper cites Choreographer: Learning and adapting skills in imagination,.

Unsupervised Skill Discovery through Skill Regions Differentiation Choreographer: Learning and adapting skills in imagination,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.987631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.647901Z digest=sha256:b1a3325a0600442a03906ac749c4aeb345c1493d7a1294f0d83a0a1e9bdb4002

Observation 57370c5e-b6cf-4ac9-a791-2a0f30a6ab51 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction,.

Unsupervised Skill Discovery through Skill Regions Differentiation Curiosity-driven exploration by self-supervised prediction,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.978763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.651425Z digest=sha256:3d82ff4ab5ee7e2c9c661aa973890f3dd302ccd28151a35636db27b313bb570c

Observation cd209605-fe50-4286-bb76-0bd9626397f5 · outbound

This paper cites Self-supervised exploration via disagreement,.

Unsupervised Skill Discovery through Skill Regions Differentiation Self-supervised exploration via disagreement,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.968809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.655593Z digest=sha256:9be038523666be5066e60b60e3f8b87037339b04a1538b6c8d149165dca6e37a

Observation 6f595c2b-5c8a-4abe-95f3-e0ea93083b26 · outbound

This paper cites Exploration by random network distillation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Exploration by random network distillation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.959226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.659560Z digest=sha256:c6db6692002ec77d465e4d8adf3d19f4e408ac7ee5a5c38e29c825a2ae994277

Observation e160b572-259a-4e80-b456-0ce9ac71b60d · outbound

This paper cites Reinforcement learn- ing with prototypical representations,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reinforcement learn- ing with prototypical representations,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.949840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.662915Z digest=sha256:d8a10cc5ca3882403a42efeda3b6b9e9a6ac67db0122f3633a5c138f3408409c

Observation 60d6b8f5-1755-43fa-be8b-c905890209f9 · outbound

This paper cites Unsupervised Skill-Discovery and Skill-Learning in Minecraft.

Unsupervised Skill Discovery through Skill Regions Differentiation Unsupervised Skill-Discovery and Skill-Learning in Minecraft

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.666160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.666160Z digest=sha256:74d632a1055c5742b0a1b3c0ebc046d71321097b36435f036f790531c613130a

Observation 0174ae71-1d5d-4b66-b325-c726bebd3b3a · outbound

This paper cites Rethinking mutual information for language conditioned skill discovery on imitation learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Rethinking mutual information for language conditioned skill discovery on imitation learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.940307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.670115Z digest=sha256:486ab42ff4e35ca9aaad085199f3fcce4fa23d3f9273313ac241fd6d533b9247

Observation 87d5dae8-0c73-47fd-adea-1a04ce1e5fd5 · outbound

This paper cites Robust policy learning via offline skill diffusion,.

Unsupervised Skill Discovery through Skill Regions Differentiation Robust policy learning via offline skill diffusion,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.930800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.673501Z digest=sha256:107cbb0a1743e9807263fb31fb84bc4bf54b0bf91c3796b6441f7d0ca38adf36

Observation 532129b3-6699-4fdc-a237-47c38bbbd6c6 · outbound

This paper cites EUCLID: Towards efficient unsupervised reinforcement learning with multi-choice dynamics model,.

Unsupervised Skill Discovery through Skill Regions Differentiation EUCLID: Towards efficient unsupervised reinforcement learning with multi-choice dynamics model,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.920765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.676484Z digest=sha256:ffa6b14e7cc41d6e5550536ce289708e9c8a75809a156ea6dc57491477b52b52

Observation 9b7851eb-dfa8-4001-93cb-37191848cf60 · outbound

This paper cites Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning.

Unsupervised Skill Discovery through Skill Regions Differentiation Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.680044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.680044Z digest=sha256:54e596b6f5228a762d0e8d9cafe35c728fad33951e174808aa7e06d6d659a0a2

Observation c0d40dd6-4f94-410f-9d1e-31652ed883fa · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice,.

Unsupervised Skill Discovery through Skill Regions Differentiation Deep reinforcement learning at the edge of the statistical precipice,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.910096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.683746Z digest=sha256:6955738ab4ba698a82e911a6aab96a17167d6de34f07d68812186cb704084911

Observation 66b480bf-5b9a-445a-93bd-0ddbd0d3441a · outbound

This paper cites DeepMind Control Suite.

Unsupervised Skill Discovery through Skill Regions Differentiation DeepMind Control Suite

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.688031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.688031Z digest=sha256:5003c7e2b4ab0a468c83f169b27d7cea3e8344db2870a7e5272c6f27d42d6d82

Observation a9f30ed3-f669-4d5f-b7d8-e7b1c2a40bed · outbound

This paper cites Continuous control with deep reinforcement learning.

Unsupervised Skill Discovery through Skill Regions Differentiation Continuous control with deep reinforcement learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.691720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.691720Z digest=sha256:68c6da14e0af06589450f6f8bc673b4d31c5bb25d3abbd07884c5d9d7b38dd11

Observation 938e136d-270b-44d4-879d-efd40bcb47a9 · outbound

This paper cites Mastering visual continuous control: Improved data-augmented reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering visual continuous control: Improved data-augmented reinforcement learning,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.900348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.695397Z digest=sha256:55f3b0b7ec051bae981fd95b3c100055a12e34d1a3a1e98d44daa9586b100c7b

Observation c14da304-2640-44aa-a0a0-67dbafd95f84 · outbound

This paper cites Mastering atari with discrete world models,.

Unsupervised Skill Discovery through Skill Regions Differentiation Mastering atari with discrete world models,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.890594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.699020Z digest=sha256:c2ddbdc32de3ec4ace090d751b95d193d02698f7d89c8c1529fc259dd4c1f0a5

Observation 05d9e23b-2958-4fc4-b1d1-809d2a638e31 · outbound

This paper cites Provably efficient rein- forcement learning with linear function approximation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Provably efficient rein- forcement learning with linear function approximation,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.880897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.702736Z digest=sha256:cda4266096a36027c626f6851b9037f0c7f3f61f70a5eb7c612abdad088e7961

Observation cf051aba-497b-4290-b167-116d05c26cdd · outbound

This paper cites Reward-free model-based reinforcement learning with linear function approximation,.

Unsupervised Skill Discovery through Skill Regions Differentiation Reward-free model-based reinforcement learning with linear function approximation,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.871448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.707318Z digest=sha256:2def72f0745b94f3a0814d8a35da4f355b56b200dfb3ae5134d517e4ddcaa2f5

Observation ae0bbdbe-c31f-4a3d-8441-e11393c504ed · outbound

This paper cites Deep variational information bottleneck,.

Unsupervised Skill Discovery through Skill Regions Differentiation Deep variational information bottleneck,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:27:00.711258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:27:00.711258Z digest=sha256:eb34c1d302a0db8e550a68563c7fc8e75bd289d9e0e5c3d9e32031dee76e5491

Observation 66625992-49c9-4b03-abe3-5d4fee27eec8 · outbound

This paper cites Dynamic bottleneck for robust self-supervised exploration,.

Unsupervised Skill Discovery through Skill Regions Differentiation Dynamic bottleneck for robust self-supervised exploration,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.854381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.715141Z digest=sha256:d81583abec54d9f9edf05bb2f9171ba62a69d71f6817d23c1c6e43c252092ca4

Observation 17e3e0cf-fcfd-4ab5-a163-7e6f6baa304e · outbound

This paper cites Provably efficient exploration in policy optimization,.

Unsupervised Skill Discovery through Skill Regions Differentiation Provably efficient exploration in policy optimization,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.844208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.718687Z digest=sha256:3cc208a9d6188b733d44a71af9ba93aa43e0874f9dbcbbe843741fa0c30e581f

Observation d95da145-63a3-467a-88a2-5c8a91764329 · outbound

This paper cites Logarithmic online regret bounds for undis- counted reinforcement learning,.

Unsupervised Skill Discovery through Skill Regions Differentiation Logarithmic online regret bounds for undis- counted reinforcement learning,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:00.832856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.722538Z digest=sha256:676ca54134f625145a146e5a8f0b34fde54eba571762c6b9a4e56c427036568f

Observation 591aa826-3252-4443-a02e-6acff41e581c · outbound

This paper cites Available: https://openreview.net/forum?id=Hkla1eHFvS.

Unsupervised Skill Discovery through Skill Regions Differentiation Available: https://openreview.net/forum?id=Hkla1eHFvS

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:27:01.008510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:27:00.640914Z digest=sha256:0055727d145451b7f3ac7237cd801a385210d68c907deeeda4509db01d126e63

Pith citing papers

No inbound Pith citation observations are available.