Pith. sign in

Paper Citation Record · LEDGER

MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.17842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.17842 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:05:27.388700Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T13:44:41.231757Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d3c1d2f-5b50-40de-90c7-60d7aa395fea · inbound

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation cites this paper.

AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T11:38:08.258353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:38:08.258353Z digest=sha256:0510aa1ff75cb77b9800a0f0d22527c883db75b0e9e1d924da387d20d1d71627

Observation 6cdbb547-8170-43c3-9e1e-9e21ee381a27 · inbound

SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text cites this paper.

SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T23:05:27.388700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:05:27.388700Z digest=sha256:835478616c17ae46cedc445109a09c6bccd786e9c73b83fc826e229c5761d4cc

Observation 88811bc7-8350-46da-a1b5-748685089f78 · inbound

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts cites this paper.

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T00:05:47.692332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:05:47.692332Z digest=sha256:b1bb6d8964b24ac65a3c852d8a92178b6ba4b7172dbaf5f7c036f8a57099317d

Observation 4ddf2e9d-3f2c-455a-888c-d41a114b55f4 · inbound

VideoASMR-Bench: Can AI-Generated ASMR Videos Fool VLMs and Humans? cites this paper.

VideoASMR-Bench: Can AI-Generated ASMR Videos Fool VLMs and Humans? MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:48:34.493010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T21:46:43.353305Z digest=sha256:6ce53b79c8ab8e25e35a83c09ee76cf62367b3739b8769938b968ef9a25d42fb

Observation 3386859a-2bfe-4c4c-ab81-38d77215093d · inbound

AVBench: Human-Aligned and Automated Evaluation Benchmark for Audio-Video Generative Models cites this paper.

AVBench: Human-Aligned and Automated Evaluation Benchmark for Audio-Video Generative Models MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:41.233058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-30T13:35:01.226818Z digest=sha256:a98f0caa6b5c03d3e704a11e5872f23cd91092d2a62dee9283837e5d2485c698