Pith. sign in

Paper Citation Record · LEDGER

High-Fidelity Audio Compression with Improved RVQGAN

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2306.06546.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.06546 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T01:50:59.050808Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bc504376-b5c5-4872-aca3-1de37ec2c946 · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding High-Fidelity Audio Compression with Improved RVQGAN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.163714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:5c00e17841ab191b43281f96687e26dc9daae613ece5cc9b89af2d0a3205b3a0

Observation c97c7d44-f0d5-45ca-8fd4-c117f5a372e6 · inbound

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns cites this paper.

SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns High-Fidelity Audio Compression with Improved RVQGAN

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-14T22:43:13.611383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:43:13.611383Z digest=sha256:0302b10653c44ffc164ac578ff31419a42f4fec1844b75ee57c0dd05eb972fa8

Observation 8da1a194-4f45-40a8-91f2-6cc16c018a92 · inbound

Woosh: A Sound Effects Foundation Model cites this paper.

Woosh: A Sound Effects Foundation Model High-Fidelity Audio Compression with Improved RVQGAN

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:15.792883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T20:51:08.144573Z digest=sha256:f19d01d388428f3914a4ce332ec87ca15bb07246ffa57d8646bc4798a3146eb9

Observation 2e20c943-ac0c-42ee-bc6b-e527083ac1ff · inbound

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs cites this paper.

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs High-Fidelity Audio Compression with Improved RVQGAN

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:18.398296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T03:13:47.971429Z digest=sha256:d0d341695885a5558cd1071ff2f0bd121564f17be8d658bed3e8cb9f8b468895

Observation d476b8f2-8481-4bbe-8e4b-97e984948d92 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:44:00.668460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T06:43:52.735211Z digest=sha256:eb4318007ccc720d15492984434b2c38eab1ede9567798b517b51aa9d68a92e4

Observation 5f42bc6e-c314-4a57-9a10-33625285b2f3 · inbound

Codec-Robust Attacks on Audio LLMs cites this paper.

Codec-Robust Attacks on Audio LLMs High-Fidelity Audio Compression with Improved RVQGAN

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:23.802002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T05:44:46.831360Z digest=sha256:bb2401006466f5713b8e13eb820d131954ec91db401cd6e12c05ccaa93c1df68

Observation 672edff9-c431-48c5-a3ec-704f628101a4 · inbound

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models cites this paper.

Inside the Latent Flow: Causal Deciphering of Attention Dynamics in Audio Separation Foundation Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:36.067248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T14:58:27.176375Z digest=sha256:e4e1d7f1faefdd7f9134d740a20c6f750aa2dcf1c7e168ded60bfd67dce636c7

Observation ad2efb5b-5feb-4db4-b3bc-8923f82e9f78 · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents High-Fidelity Audio Compression with Improved RVQGAN

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.521886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:8d13da4e9c127ec6e57d9c5cea58126099ce1af95b77577efb631f8b94c21016

Observation c0923909-5374-4f44-94ef-df9eedf159f7 · inbound

NAC: Neural Action Codec for Vision-Language-Action Models cites this paper.

NAC: Neural Action Codec for Vision-Language-Action Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:59:37.904034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T13:59:53.484305Z digest=sha256:469ad5eee0def38771d09fa075483df35f41dc5f98cb184ac78ab1865345e6c7

Observation 5d2969a1-f330-4ccb-8171-eb5121f4d6cd · inbound

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models cites this paper.

HybridCodec: Modeling Discrete and Continuous Representations for Efficient Speech Language Models High-Fidelity Audio Compression with Improved RVQGAN

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:02:56.247716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T00:56:43.991936Z digest=sha256:617ffdb339c7b8b5f071960bc3a4944843f3789f571c0d7a921a0e12350fa292

Observation b520bed0-bfc9-4c40-a6f7-2a2ee91cd0bf · inbound

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts cites this paper.

ITGPT: A Transformer Based Architecture for the Generation of Dance Dance Revolution and In the Groove Charts High-Fidelity Audio Compression with Improved RVQGAN

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:30.515762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:23:30.515762Z digest=sha256:f0af2c6c12510678fd1200ad41d04a35ffd191ecfe6bfec2e539859a91768d13

Observation 753f55db-093b-4210-a661-34c9eb1e7920 · inbound

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness cites this paper.

Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness High-Fidelity Audio Compression with Improved RVQGAN

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T08:25:28.089701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:25:28.089701Z digest=sha256:9643f95754c166107903d3a3ba0671b64d86adcba22acbcefd63efc704c59c04

Observation ffac6167-9e67-498e-ba85-c20864d82cf3 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T10:35:00.630413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T10:35:00.630413Z digest=sha256:25e5022497e1976c70810cacb9e0430dc929bb90226f5b709a4c29bf6fb918be

Observation b693b372-f65d-4222-9903-f5161639ba97 · inbound

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation cites this paper.

OmniVAE: An Audio-Video VAE with Cross-Modal Alignment for Joint Generation High-Fidelity Audio Compression with Improved RVQGAN

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T01:50:59.050808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:50:59.050808Z digest=sha256:d063a77f504d803a74893a97b543b091f12edf328a7049afbce021e20c3770f6