Pith. sign in

Paper Citation Record · LEDGER

Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2404.09956.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.09956 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:52:01.341231Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.588067Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 660d8a96-811c-40bf-88fb-5779fd49f8e0 · inbound

Movie Gen: A Cast of Media Foundation Models cites this paper.

Movie Gen: A Cast of Media Foundation Models Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:25.186220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T14:16:18.521699Z digest=sha256:c4096e7aaf99cb7451ca1cfd33bd463b54ba8d948be1e2b01f3eef2aaa430bc8

Observation abc45090-463f-488d-aa53-1b2bcc510a71 · inbound

Generative Audio Language Modeling with Continuous-valued Tokens and Masked Next-Token Prediction cites this paper.

Generative Audio Language Modeling with Continuous-valued Tokens and Masked Next-Token Prediction Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T17:52:01.341231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:52:01.341231Z digest=sha256:0958d4a1bbc3993ebc54387c61a649c17e1eba4c93bb72044f9dd8e044228acb

Observation 0e9d28d4-4408-4b6d-9bbd-29e4bd1a5891 · inbound

JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment cites this paper.

JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T13:14:37.557921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:14:37.557921Z digest=sha256:cfcb75597ee5e3eeff9ecbf22de30eac722c0f12c5a4fac8773ce82569901438

Observation fae8cc7b-3739-47c2-811e-f57b1772e517 · inbound

SemanticAudio: Audio Generation and Editing in Semantic Space cites this paper.

SemanticAudio: Audio Generation and Editing in Semantic Space Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T07:06:33.285917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:06:33.285917Z digest=sha256:7c88e4ffb3faab894d97fec18fea43f1fd4687f63a4aa8b06a33b838b14d5270

Observation 785ad587-19b5-4d6a-ad29-e492ebf26b8d · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Reference 99

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.589338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:95206359552bb5004bbaede3bce7aa992a9413038bd4ede97eecea2f21a9d4b8

Observation 336a520b-525f-4076-a526-ebe9ed6dea0c · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Reference 99

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:f870e95337539fd672316dcacb91aaa027959fa50981a2fcd3f0e91d68786090