Pith. sign in

Paper Citation Record · LEDGER

Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2401.08567.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.08567 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T20:40:32.691742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T08:17:36.053151Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2e4a100a-d1a2-4222-b52e-8e0c8a80e380 · inbound

ResidualDroppath: Enhancing Feature Reuse over Residual Connections cites this paper.

ResidualDroppath: Enhancing Feature Reuse over Residual Connections Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-12T20:40:32.691742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:40:32.691742Z digest=sha256:7db56ec8351edb10c3f064c30725941e3549ab3e2d2f7c79398e040bd8b6c4fa

Observation 5eaadca9-61ad-4a4d-8648-2a6505454f64 · inbound

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models cites this paper.

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data

Reference 40

Resolution
malformed identifier
arxiv_id, observed 2026-05-16T08:17:36.055156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-16T08:17:29.924860Z digest=sha256:6dc3a5ad257c98ddfad8cee11c598c6fa0eca8aaa441d4fc001eaf169d229866

Observation 9f526d3a-ddaf-4edb-8467-02536b0f75dd · inbound

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models cites this paper.

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data

Reference 40

Resolution
malformed identifier
no resolver link, observed 2026-08-03T05:34:31.783227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:34:31.783227Z digest=sha256:c0f234b3d5cbaa9bdd530706b048e82fccf9a0354ba226b4ea2b0c406ab18642

Observation 9fe4460e-3293-4340-8d8e-63223e371fe2 · inbound

Anisotropic Modality Align cites this paper.

Anisotropic Modality Align Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:55:56.205648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T02:08:04.686087Z digest=sha256:76fffd16763227b802b2aaf01fc38584f4495a2cb25dbdbeb324f11273ece957

Observation 19c991cc-536b-47da-9036-bc94ea65deb0 · inbound

Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioning cites this paper.

Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioning Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:17.384647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:17.384647Z digest=sha256:ac246ccf9585db1ceba05c8164bd97610562d0d9a0c461fb17b6abfad2de9578