Pith. sign in

Paper Citation Record · LEDGER

Dynamics-Aware Unsupervised Discovery of Skills

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:1907.01657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.01657 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:11:08.934687Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:39:34.298932Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5acb6b12-c7c7-4c9e-accb-c3b05f64da49 · inbound

Is Conditional Generative Modeling all you need for Decision-Making? cites this paper.

Is Conditional Generative Modeling all you need for Decision-Making? Dynamics-Aware Unsupervised Discovery of Skills

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:35:10.807333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T15:35:10.593969Z digest=sha256:a9f6d1d3cddba357d24227a4c8acaf437bb91c7bf2d1eada861696f751251ae4

Observation 7f974e1c-aa56-4a94-9545-65687568ef72 · inbound

Training RL Agents for Multi-Objective Network Defense Tasks cites this paper.

Training RL Agents for Multi-Objective Network Defense Tasks Dynamics-Aware Unsupervised Discovery of Skills

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:11:08.934687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:11:08.934687Z digest=sha256:15ca62c6c151a4827a8bf096e59337d04b7cf024c4b11349931acf52941148f9

Observation e03d3241-ec77-42e9-b744-9191bd314146 · inbound

RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming cites this paper.

RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming Dynamics-Aware Unsupervised Discovery of Skills

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:20.279140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:20.279140Z digest=sha256:1d633d2c99c0f20367aba34e7317806f623f537402b94d0427ee741225d3e3c1

Observation 6c34781e-890b-4cc6-b75c-9f2be1a45cb6 · inbound

Epistemically-guided forward-backward exploration cites this paper.

Epistemically-guided forward-backward exploration Dynamics-Aware Unsupervised Discovery of Skills

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:32:34.807117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:32:34.807117Z digest=sha256:f448324b59c458c078ee672b6293c2a99461d734db13fbb2d976c6e0c20d9186

Observation 15e5ecc8-53f5-4941-bf8a-32a8a5e781e0 · inbound

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning cites this paper.

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T15:56:59.109038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:56:59.109038Z digest=sha256:f109a8400d581d4d04d0528fd8eab06a45e3f57376a3cbe38e225ca617820622

Observation 278294fc-c0de-4947-8e82-edd4710c55d0 · inbound

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs cites this paper.

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs Dynamics-Aware Unsupervised Discovery of Skills

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:34.996582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:17:34.996582Z digest=sha256:48406500a4badf5fae8c3d78e64c8278ea4d3a1ab49f83fd152c3b51b639435c

Observation f91f9964-d241-4916-963c-270cb71cd320 · inbound

Hierarchical Behaviour Spaces cites this paper.

Hierarchical Behaviour Spaces Dynamics-Aware Unsupervised Discovery of Skills

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.758712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T03:33:48.527621Z digest=sha256:6f8477d5396ec0b41e4bdc0bf83f31e2fa53dd8a8f1e0a7569cd3ccab12fa187

Observation f3455a21-b401-4303-8cf9-5db592cdf331 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Dynamics-Aware Unsupervised Discovery of Skills

Reference 165

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:58.267267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:465a4fe749e6fa04c2325fda35929a4c87a26e867521c4f2911232b340898dfd

Observation e924b165-9681-4871-ab40-5a8760a8dc21 · inbound

Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization cites this paper.

Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization Dynamics-Aware Unsupervised Discovery of Skills

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:46:10.749389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T13:52:43.033237Z digest=sha256:ddf9e479f4a1c7a809aa4bc1ccbb0ab90c189c92454ad6b38692136ec1c15226

Observation 686c3abf-358b-43e8-93d9-8e92095cf7e7 · inbound

Delay-Empowered Causal Hierarchical Reinforcement Learning cites this paper.

Delay-Empowered Causal Hierarchical Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:47:21.205769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T05:46:51.659283Z digest=sha256:f229252dd64bee7fe785bb8acf057aa8e50f0d7bc5661cd7b4a60c567b1d6cc4

Observation 1f4140f7-9c74-463d-bec6-c24b759d7fe7 · inbound

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning cites this paper.

Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning Dynamics-Aware Unsupervised Discovery of Skills

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:26.372151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T18:37:48.259353Z digest=sha256:00aa10a2d2ae98ba80299f49174cdb37fd4cd5faeadd7bd83481c225a03c58ac

Observation 03c1cede-f8a5-4dd2-8a67-1c86337e9095 · inbound

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling cites this paper.

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling Dynamics-Aware Unsupervised Discovery of Skills

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:07:34.177073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T16:07:28.361043Z digest=sha256:6bc7cfa8f4d2ad98620da74d3ef9161805dc75b9f52037c643e6077a7c848bd6

Observation 10b55895-2a64-4cd9-b449-f9a9328fd0fc · inbound

Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory cites this paper.

Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory Dynamics-Aware Unsupervised Discovery of Skills

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:39:34.301049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T16:53:53.441274Z digest=sha256:3c13eb5d241217b578a48108481004bea7c5fea065e0badcea9db5e56947159c

Observation b7d65f84-7599-455f-a48e-20e67a2266d8 · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering Dynamics-Aware Unsupervised Discovery of Skills

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.206134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:8e3614bf23085abbcf71f4ff0773b988242939ee0c830e6ff27b4d780f510b63

Observation 534371e1-d8ba-4b70-9bdd-0db6848e22e7 · inbound

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies cites this paper.

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies Dynamics-Aware Unsupervised Discovery of Skills

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T06:32:52.458556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:32:52.458556Z digest=sha256:2854aef688c50dbb801ee214a14527dfad8b2513b0197c318a740138ce56de98

Observation 1d53a89b-01a5-44af-929d-7b3fbc749eb3 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Dynamics-Aware Unsupervised Discovery of Skills

Reference 229

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:35.216572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:35.216572Z digest=sha256:5cbc75f7b642aa2fd4b4ec54398fb8707f60404008e4d0ab5b6dc21ca4713b1e