Pith. sign in

Paper Citation Record · LEDGER

Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.05298.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.05298 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:10:56.224674Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:48:12.303594Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 08f35b59-de99-4771-8b28-d71248dd34db · inbound

NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference cites this paper.

NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:10:56.224674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:10:56.224674Z digest=sha256:7f106bdc9075cd0e5802d33a929d6acc32f5b33bfb4f0f244674b2cbfd9e256c

Observation 33c1a620-8ab2-404d-b00e-c8e949a92a2d · inbound

Is GAN Necessary for Mel-Spectrogram-based Neural Vocoder? cites this paper.

Is GAN Necessary for Mel-Spectrogram-based Neural Vocoder? Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T21:58:47.274633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:58:47.274633Z digest=sha256:ab26969ae2ae95ab51d6e481043f64df251b8cda76972ac2e341e1cbf204a6a8

Observation 1886f8f4-88fc-4818-9a5b-9a573165cfbb · inbound

Representing Speech Through Autoregressive Prediction of Cochlear Tokens cites this paper.

Representing Speech Through Autoregressive Prediction of Cochlear Tokens Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T19:54:58.582582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:54:58.582582Z digest=sha256:f0bb1116a21adcdfb1048b1206f2f6abb7bb001361e9c2804b7adaf9ec78bed1

Observation 64b8a3d7-a39a-4739-bfab-3d83e9ff219c · inbound

Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation cites this paper.

Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T18:47:43.072454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T18:47:32.003545Z digest=sha256:0fc3f9734b2849cf8c4d5bc42450f2910e34996e02b7d5d36645efdc0f4aabf9

Observation 40958024-9c6f-47c9-b709-e4f5379d2e60 · inbound

Ultra-Low-Bitrate Mel-Spectrogram-based Neural Speech Coding with Flow-Matching-based Refinement and Vocoding-driven Reconstruction cites this paper.

Ultra-Low-Bitrate Mel-Spectrogram-based Neural Speech Coding with Flow-Matching-based Refinement and Vocoding-driven Reconstruction Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:53:55.890818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T19:45:01.705167Z digest=sha256:3ebcfe43c5b895e0b060cfab64ea8e246a8ccfe8aadf2dd6d4c9a10e6a9851aa

Observation 8bb0167e-4a76-496d-acfa-cef0bee442fd · inbound

Benchmarking Neural Speech Compression from a Rate-Distortion Perspective cites this paper.

Benchmarking Neural Speech Compression from a Rate-Distortion Perspective Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-03T12:48:12.305091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T08:43:35.279033Z digest=sha256:e1be5f4c35ad1ac03253d4381147e3cf151f9750079b102c4ecb18e108a27850