Pith. sign in

Paper Citation Record · LEDGER

A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2408.05804.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.05804 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:41:37.321403Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T16:54:57.940762Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8972b79e-4f56-4c4d-879d-e3ff63f6d453 · inbound

Efficient Skill Discovery via Regret-Aware Optimization cites this paper.

Efficient Skill Discovery via Regret-Aware Optimization A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:37.321403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:37.321403Z digest=sha256:468484ca4c7d0c5c9738a9109feaa276719ad2f0e0c68d018b18a65286058368

Observation 84216d40-7969-43ad-970a-15f13301d793 · inbound

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning cites this paper.

Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:56:56.081279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:56:56.081279Z digest=sha256:4e1273e93613e1d52c159be5c54428b1b9901683473576a7cd0e90bb35e77c6e

Observation 10ce83fd-e7cf-4ee8-8715-835089961f21 · inbound

Equivariant Goal Conditioned Contrastive Reinforcement Learning cites this paper.

Equivariant Goal Conditioned Contrastive Reinforcement Learning A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:24.562116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:24.562116Z digest=sha256:232f430b23efe1501a33aa27b38dc496ae32a911e1c3e1a16476f5e67c821d9c

Observation 0d1ae859-7443-4b2b-bf64-c66a6dfdb688 · inbound

Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration cites this paper.

Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T17:49:13.985316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:49:13.985316Z digest=sha256:ce206fec11ca28bd97b82db9fd6f39a9618c402d688b8978be4b9febbe6e7b0d

Observation 44d2445e-c450-48cf-9206-fbc706025af1 · inbound

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry cites this paper.

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-13T14:05:26.303000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:05:26.303000Z digest=sha256:c6626e2be930fa4e15348fa517044d19c17f13785751a6c0717459cf97a0d570

Observation 3063058f-6123-4061-9122-c9530a1e4e0a · inbound

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning cites this paper.

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:20:58.628509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:12:32.839681Z digest=sha256:4f82bbebd960de5cdd059ba171c4ca95e38dfd12b4ba04bf43f6f373cd481894

Observation c59c9500-ea23-4fd3-933a-c610607126f1 · inbound

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation cites this paper.

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:12:55.194473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:12:07.920183Z digest=sha256:002da5b90ea90be9186618311e752f8cb6bf255eeb346785d5a4a804723479d1

Observation 2e852c91-7f39-467e-9b2f-b02d6c46339a · inbound

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling cites this paper.

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 284

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:08:59.676998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T03:05:36.871497Z digest=sha256:225c4c2c7691f46fe0aaaaaab07efbd091c8479c9bae4ee8a07cfb959805adde

Observation 7a41a0a8-e1ca-459f-91a9-02096bbaefb2 · inbound

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation cites this paper.

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:41:08.508353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T05:37:26.290308Z digest=sha256:39e883e3402781710c7d8f2ddd5b0c44adede98e65918aad10fb7f132d05460a

Observation 4aa81f55-ab8b-455a-9022-22a0aa3721ff · inbound

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation cites this paper.

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:54:57.942185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T16:54:39.269896Z digest=sha256:c3c4f6f8a8e98538d728d6985bd98d9806b18d969fd050d4a7ec50a575b23abb