Pith. sign in

Paper Citation Record · LEDGER

RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2111.05011.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2111.05011 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:43:40.203915Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b0d1201b-80cf-4fc9-8afc-5498299bf0bb · inbound

Aligner-Guided Training Paradigm: Advancing Text-to-Speech Models with Aligner Guided Duration cites this paper.

Aligner-Guided Training Paradigm: Advancing Text-to-Speech Models with Aligner Guided Duration RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T18:16:26.951656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:16:26.951656Z digest=sha256:2a31f679c40436bd9715703856048eff1823e98a58f4b8929b221ee772021b78

Observation a5167604-be12-4ec3-8600-53b7c9ad8409 · inbound

LatentSpeech: Latent Diffusion for Text-To-Speech Generation cites this paper.

LatentSpeech: Latent Diffusion for Text-To-Speech Generation RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T18:16:12.925172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:16:12.925172Z digest=sha256:11894ba736c487245eeea8d30d343bfc2a10b078a4ee7ac0fbf4d57ac4dbc8c4

Observation 0f5387d0-b9ca-4fc2-95b3-38875f900125 · inbound

Hidden Echoes Survive Training in Audio To Audio Generative Instrument Models cites this paper.

Hidden Echoes Survive Training in Audio To Audio Generative Instrument Models RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:50:02.164251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:50:02.164251Z digest=sha256:90f3feecef5f449f7c21c0261d649c588e6e964bfb3b20c9779e1da1298a27ba

Observation 811671bf-a42e-4885-8822-c494bf1a4661 · inbound

Drivetrain simulation using variational autoencoders cites this paper.

Drivetrain simulation using variational autoencoders RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:33.398716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T04:29:41.412588Z digest=sha256:6f4610ac8e7afaeea8bf8d9a680a36e72323bd9b911ca7ade79029f45c5dfe54

Observation fd93cc21-d578-4c87-805d-050601a9c620 · inbound

OmniAudio: Generating Spatial Audio from 360-Degree Video cites this paper.

OmniAudio: Generating Spatial Audio from 360-Degree Video RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:43:40.203915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:43:40.203915Z digest=sha256:0e595c6f2e1c8873598bc903acdb7653aaf8f55343a0d125fe89a9367ff56832

Observation 9a4e02bf-85c5-4298-8f6f-84f6c2163f61 · inbound

Learning to Upsample and Upmix Audio in the Latent Domain cites this paper.

Learning to Upsample and Upmix Audio in the Latent Domain RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:44.883723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:05:44.883723Z digest=sha256:abce981cdc35373ed9f65f555d6c26097f632bcda8e90b0cdb0122c9f2bafbca

Observation 034f9c55-3ae4-489c-bcdc-664b54371171 · inbound

ANIRA: An Architecture for Neural Network Inference in Real-Time Audio Applications cites this paper.

ANIRA: An Architecture for Neural Network Inference in Real-Time Audio Applications RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:48:30.536237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:48:30.536237Z digest=sha256:9539a8f666f8b75429dd606ff66e532f21a41211bf5e5d1e7f9509a0e222c729

Observation c3317019-c7e1-4ca6-9705-d7e113591a1d · inbound

Two Sonification Methods for the MindCube cites this paper.

Two Sonification Methods for the MindCube RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:26:43.953857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:26:43.953857Z digest=sha256:fb030ec66501aa7b56e6066605f547c6494fcfa16a2f72a43f52a9afbb43e637

Observation 0a9287ce-e4a3-47d7-a1ea-b9f7b7db5fe8 · inbound

Workflow-Based Evaluation of Music Generation Systems cites this paper.

Workflow-Based Evaluation of Music Generation Systems RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:03.295053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:46:03.295053Z digest=sha256:43612be6aec51c85dbd0c588468d5a5a6ba75cc1e0fe8b61fea0f0048f30c6ea

Observation 82963036-d74f-47ff-818e-47f5f3538017 · inbound

MusGO: A Community-Driven Framework For Assessing Openness in Music-Generative AI cites this paper.

MusGO: A Community-Driven Framework For Assessing Openness in Music-Generative AI RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:09:33.775638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:09:33.775638Z digest=sha256:badde6ec4aa12420f4982c3f4fb4089d7bc2cb07319c3d9ca6ff7b1f0ff270ca

Observation e2c2c975-d8df-4d5b-a744-a931dad21161 · inbound

Latent Fourier Transform cites this paper.

Latent Fourier Transform RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:21:07.037225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T03:45:07.892234Z digest=sha256:6f4bb6d2a3933e6d74c9c4fb715f1f5e60cc8f454c19ed211453005a2547e326

Observation 103196e1-b5ee-4830-8e46-3370a3f0ce64 · inbound

Opening the Design Space: Two Years of Performance with Intelligent Musical Instruments cites this paper.

Opening the Design Space: Two Years of Performance with Intelligent Musical Instruments RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:13.486349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T05:24:28.773536Z digest=sha256:bb7d8ecb8634523a8ccf789d23cc64dc05ff7ad6279dd0b5e284112b491f9a9f

Observation 7fb4adcd-3a3c-4563-89c8-7ad2ec66b7e3 · inbound

Hu\'i S\`u: Co-constructing a Dual Feedback Apparatus cites this paper.

Hu\'i S\`u: Co-constructing a Dual Feedback Apparatus RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-09T02:49:42.758802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T14:37:42.783784Z digest=sha256:dcb6abd4dafed16d97158f5d024c2b4d5585b61512666fd2d228d5b1c7054fd5

Observation b133ed4d-858f-4fb6-b20c-43fbca023065 · inbound

Remix the Timbre: Diffusion-Based Style Transfer Across Polyphonic Stems cites this paper.

Remix the Timbre: Diffusion-Based Style Transfer Across Polyphonic Stems RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:46:27.992987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T04:58:26.634355Z digest=sha256:3d323ae723091ab7cfe1c28ae913c9fff7500a8d1806a98940780b6a01e633fb

Observation ec707f93-ec5c-4e6d-a098-720c679b1ebc · inbound

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators cites this paper.

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T03:25:58.969324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T03:24:50.604019Z digest=sha256:b3aa9689012af13a3bf07625e0f5f6ebe006a43515b6853795f549fa22cd74cf

Observation 4eebc060-3237-4bfe-b8bf-90a86a8f6292 · inbound

DEMON: Diffusion Engine for Musical Orchestrated Noise cites this paper.

DEMON: Diffusion Engine for Musical Orchestrated Noise RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:03:17.319218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T09:54:30.079404Z digest=sha256:3dd47b891a20a421f32a3096384488b7c0b985061909883798b8c238d0381f8c

Observation 2c92f274-b3f4-4390-9323-1e6fc2216daf · inbound

FXplorer: A Map-Based Interface for Exploratory Audio Effect Design cites this paper.

FXplorer: A Map-Based Interface for Exploratory Audio Effect Design RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:07:26.641255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T19:08:21.640049Z digest=sha256:198b66bfd09bd3aec8b004b54e3b0c8c7c52aedb06891fbb839b773bacff8fd4

Observation ac75d854-52a2-4660-a6a8-43dca1258c85 · inbound

Structural Bottlenecks on Frequency Representation in End-to-End Audio Models cites this paper.

Structural Bottlenecks on Frequency Representation in End-to-End Audio Models RAVE: A variational autoencoder for fast and high-quality neural audio synthesis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.385905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-10T05:37:12.972599Z digest=sha256:c63c7d3d8276e346f710af67eb756e8ffa88ecf3a74d516e828860f4a36d643f