Pith. sign in

Paper Citation Record · LEDGER

Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2405.15613.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.15613 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:47:32.867409Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:17:23.968320Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 15b3e917-71af-40ce-ac06-15dce7695f8d · inbound

LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics cites this paper.

LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:23:00.067610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-16T07:22:59.854042Z digest=sha256:d0bd5e015d36753af1c07543e0acaf10122ca8b405964e81454e63c1a1eb2cc3

Observation 6531babd-6eb0-4ada-812c-79045f6e0031 · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:08:37.130097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T14:08:36.801359Z digest=sha256:d928f3b6afed1da49fd66e2fa58359beb8bb7489468cec23e1d22bfd12c2cf84

Observation 2c059e89-b1c9-4c53-9b29-e88f08179e47 · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T19:47:32.867409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:47:32.867409Z digest=sha256:c4bfc7806ea87f3f9880e7e440a0b4c129b661b59ace7ea9773c473d6fd2694a

Observation b1975d65-e87a-4506-a615-fca3d7a7a082 · inbound

SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models cites this paper.

SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:23:23.743697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T20:21:47.430865Z digest=sha256:f9705ebf4f8c6d4aeeea8ecf111536a8b90a1d9b18364ed8aaa7bd2b34cd1ed0

Observation bc0ab429-8975-48f9-a4bb-b7f96a9abf09 · inbound

MegaStyle: Constructing Diverse and Scalable Style Dataset via Consistent Text-to-Image Style Mapping cites this paper.

MegaStyle: Constructing Diverse and Scalable Style Dataset via Consistent Text-to-Image Style Mapping Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:51:11.129675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:54:25.679094Z digest=sha256:58123fd4361b3532290324b85d07dce5930f5844a1d8b48a37cd3bb4c6084f2c

Observation 9cb17e4a-9f86-40dd-af5a-c3f964cc6b53 · inbound

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer cites this paper.

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:46:49.103971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T04:15:20.045060Z digest=sha256:1c27b7d071a4e674614f0c31672756924dd52bb1231e2372654f918cd103da25

Observation 45ff0784-87e2-48e5-b3e7-663d81821d96 · inbound

DataComp-VLM: Improved Open Datasets for Vision-Language Models cites this paper.

DataComp-VLM: Improved Open Datasets for Vision-Language Models Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 292

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:47.629284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T01:16:16.834861Z digest=sha256:8cfbdef20e048ab83aab431285cc8e8b71ae3fbd584fe52005580995bdaaa06a

Observation 3a1673dd-3636-4679-b09d-3b1513be82a2 · inbound

DataComp-VLM: Improved Open Datasets for Vision-Language Models cites this paper.

DataComp-VLM: Improved Open Datasets for Vision-Language Models Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach

Reference 292

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:17:23.969751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-02T21:10:10.548489Z digest=sha256:3f6e9dbb2ab19d3e46e5f861080b0371d13a99f0032c47337b4efede1f68412f