Pith. sign in

Paper Citation Record · LEDGER

High-Fidelity Audio Compression with Improved RVQGAN

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2306.06546.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.06546 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:45:14.517623Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

24
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cbab5824-3108-4643-98fc-d56a7c703027 · inbound

Compression of Higher Order Ambisonics with Multichannel RVQGAN cites this paper.

Compression of Higher Order Ambisonics with Multichannel RVQGAN High-Fidelity Audio Compression with Improved RVQGAN

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:03:55.346910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:03:55.346910Z digest=sha256:3cfeef6d88a78f87611fd369eaeb19e88fbb275efdaa9fd23be765146058d5fc

Observation eb938b15-2c87-467d-814e-7b3490621349 · inbound

Watermarking Training Data of Music Generation Models cites this paper.

Watermarking Training Data of Music Generation Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T17:49:42.586137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:49:42.586137Z digest=sha256:5ebfd740acf353c30efe3e2d9bc6c0c933e3d19b3fa2a630f4d9f876a244a415

Observation af746ae4-eecd-401a-8723-5f5a39208e55 · inbound

HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution cites this paper.

HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution High-Fidelity Audio Compression with Improved RVQGAN

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:35.358913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:35.358913Z digest=sha256:96b656e84b4cc22f379ade7bcd0e6b171b50f4e28f24212ccf6425b9024f8c62

Observation 912afb78-bb55-48e2-92fd-f0e6fd2bec8a · inbound

Aliasing Reduction in Neural Amp Modeling by Smoothing Activations cites this paper.

Aliasing Reduction in Neural Amp Modeling by Smoothing Activations High-Fidelity Audio Compression with Improved RVQGAN

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:45:14.517623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:45:14.517623Z digest=sha256:537d38d270b8fcdbd78bbdb289470c6160b062715037f27e2df883a212dd788c

Observation d905ae10-cab3-4834-910b-4341474c8dc1 · inbound

Toward a Sparse and Interpretable Audio Codec cites this paper.

Toward a Sparse and Interpretable Audio Codec High-Fidelity Audio Compression with Improved RVQGAN

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:02:53.134621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:02:53.134621Z digest=sha256:ae69667c5c4a61cf4639a4064ba48cef92b71fe0e986ea0f12626ae4ac05d7c2

Observation 3e270137-8315-4fb3-9ec5-c3805b530c73 · inbound

AI-Generated Song Detection via Lyrics Transcripts cites this paper.

AI-Generated Song Detection via Lyrics Transcripts High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:21:30.071947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:21:30.071947Z digest=sha256:66c080bb53c0ebf1158069ef8917c15d14a34954e1449ef5cabbbd7d4298fd27

Observation 72e08b84-24df-4c1c-8c9d-fa09d208c4c9 · inbound

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine cites this paper.

Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine High-Fidelity Audio Compression with Improved RVQGAN

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:34.455589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:47:34.455589Z digest=sha256:e29ab740afaf63d676583b8f602f3b7dd6834af1d8601c70ec903073ff0aa8f5

Observation 2798e571-ad33-4e5a-ada7-a1d5ba7f446d · inbound

Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning cites this paper.

Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning High-Fidelity Audio Compression with Improved RVQGAN

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:14:46.240127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:14:46.240127Z digest=sha256:b59e5a2863628bdc4a96543ad86e9be8a4ebc3382289881148d80b9bf786adda

Observation 0ea89150-798c-40b1-9ff6-63b4d87fc4ef · inbound

Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission cites this paper.

Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission High-Fidelity Audio Compression with Improved RVQGAN

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T11:29:01.971972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:29:01.971972Z digest=sha256:81b646e4d4e5701de270dce3522dcc0abd546c3184cd9cbdf334f258b342212f

Observation bc504376-b5c5-4872-aca3-1de37ec2c946 · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding High-Fidelity Audio Compression with Improved RVQGAN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.163714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:d3073fafbcbdf1e732a13205368939bb48e15754e3322b3125cea438d03e620b

Observation c97c7d44-f0d5-45ca-8fd4-c117f5a372e6 · inbound

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns cites this paper.

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns High-Fidelity Audio Compression with Improved RVQGAN

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-14T22:43:13.611383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:43:13.611383Z digest=sha256:bd2b565d7f5e9721d6a164d4843a086dfd906d2c21a769dd4432b8e00a2c46d0

Observation 8da1a194-4f45-40a8-91f2-6cc16c018a92 · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model High-Fidelity Audio Compression with Improved RVQGAN

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.792883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:2a596a4b4238f5b6993b8a4db3ff615ca830d39268025c945ed209d0e155426a

Observation 2e20c943-ac0c-42ee-bc6b-e527083ac1ff · inbound

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs cites this paper.

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.398296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T03:13:47.971429Z digest=sha256:2a2c4ae085cc50658efdaa0a153eabd7ffbf2eb051f1f14e3c976786b77d33d9

Observation d476b8f2-8481-4bbe-8e4b-97e984948d92 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:44:00.668460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T06:43:52.735211Z digest=sha256:35559f4b44374870d8df9a921d92cdbb1edaafc80742d91e30985d50a19a279d

Observation 5f42bc6e-c314-4a57-9a10-33625285b2f3 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:23.802002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-25T05:44:46.831360Z digest=sha256:77086b6d6857ad29076e72c0af802b21650c5070d6a831eae1946ce7b4a01178

Observation 672edff9-c431-48c5-a3ec-704f628101a4 · inbound

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models cites this paper.

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:36.067248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T14:58:27.176375Z digest=sha256:bcc9a270d3ef22e595a40e7b39b2a16ff180e73273960e98a8d686b15f88afd1

Observation ad2efb5b-5feb-4db4-b3bc-8923f82e9f78 · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents High-Fidelity Audio Compression with Improved RVQGAN

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.521886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:f73c6a5d68a559e05a844ac94e3997da5eefb3c06b7ca5964cfafb19dc1f2f9c

Observation c0923909-5374-4f44-94ef-df9eedf159f7 · inbound

NAC: Neural Action Codec for Vision-Language-Action Models cites this paper.

NAC: Neural Action Codec for Vision-Language-Action Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:59:37.904034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T13:59:53.484305Z digest=sha256:c820df46628c3cb45dee3046ed17668940ee406e573bb73d5f103c60ea965765

Observation 5d2969a1-f330-4ccb-8171-eb5121f4d6cd · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.247716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:9bf63c055dae7f12b34dea6ede45de5e4b27f2ea8a759bf6a337454d6d583fda

Observation b520bed0-bfc9-4c40-a6f7-2a2ee91cd0bf · inbound

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts cites this paper.

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts High-Fidelity Audio Compression with Improved RVQGAN

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:30.515762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:23:30.515762Z digest=sha256:7bd14a9830d8a9e53f9d2b080687c3ec43b690d6e3ff12cd066b8a378c443722

Observation 753f55db-093b-4210-a661-34c9eb1e7920 · inbound

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness cites this paper.

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness High-Fidelity Audio Compression with Improved RVQGAN

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T08:25:28.089701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:25:28.089701Z digest=sha256:38cf79860c3077242cfd054cb7315ba4c2dae14572796cf29467990f40c5e5b7

Observation ffac6167-9e67-498e-ba85-c20864d82cf3 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T10:35:00.630413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T10:35:00.630413Z digest=sha256:2a8d858d0d8e9dd32f182ceb15897e9ec55c822d33950779d23ae84cca9e6081

Observation b693b372-f65d-4222-9903-f5161639ba97 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T01:50:59.050808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:50:59.050808Z digest=sha256:b032b1bbd86809ffdc71c52e6f1ddf5b27ce086b6175bd93f9605d43ade3eb70

Observation 6998a798-10b3-4f4b-b2bb-868dab788b76 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks High-Fidelity Audio Compression with Improved RVQGAN

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T16:29:24.694740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:29:24.694740Z digest=sha256:9b6b8776520effce5bc21b6010e29f1ad24a3b26f24bf496d773b83db6ca2b90

Observation 59066311-494b-4bd7-b822-fd5f4f5b1025 · inbound

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks cites this paper.

SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks High-Fidelity Audio Compression with Improved RVQGAN

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:45.767421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:45.767421Z digest=sha256:c42f8a334909a40775edf8091d7cccde6558bc01bbc2479851ba6cd9ab0f9abc

Observation 613f7c29-7e12-4080-b576-24f4866498b2 · inbound

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs cites this paper.

On the Geometry of Music Bandwidth Extension in Latent Spaces of Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T14:08:33.591322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:08:33.591322Z digest=sha256:67936cc842334b8d2d954bafc74f018468f268bff94c7818ea37341eb43e1330