Pith. sign in

Paper Citation Record · LEDGER

Rate-Aware Learned Speech Compression

As of 13 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2501.11999.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.11999 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:41:19.298523Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 89dd27ab-05e1-4892-853f-87457387c87e · outbound

This paper cites Definition of the opus audio codec,.

Rate-Aware Learned Speech Compression Definition of the opus audio codec,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.192357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.192357Z digest=sha256:ccdf00773f3398c7cedcd61058f0e31a8e6490324538ec2a2901def7ffa02e9a

Observation 22304633-4246-4270-a684-e90df86c19f7 · outbound

This paper cites Overview of the evs codec architecture,.

Rate-Aware Learned Speech Compression Overview of the evs codec architecture,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.657670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.198238Z digest=sha256:c59b036f68fbb8cf248ea069d33ca435287a4a7e987158ea64619bd556b37460

Observation 80acd235-df74-4ede-bcfc-d6e70f13a565 · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

Rate-Aware Learned Speech Compression Soundstream: An end-to-end neural audio codec,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.639745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.203578Z digest=sha256:38ae5435da7d5f0b2a3f368df038724519ab30fc200cb45f8241ee824f44ec64

Observation bfe1fc2a-de89-48a8-a7b0-7b9e9a7722b4 · outbound

This paper cites High Fidelity Neural Audio Compression.

Rate-Aware Learned Speech Compression High Fidelity Neural Audio Compression

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.208986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.208986Z digest=sha256:4475daff7ca29d560db61d1aae24e1d558d89ee28a2bb5247e5fe9e83196ca66

Observation 77f6a551-29f8-4e24-bfb7-a67cb887a95a · outbound

This paper cites Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,.

Rate-Aware Learned Speech Compression Funcodec: A fundamental, reproducible and integrable open-source toolkit for neural speech codec,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.619889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.215105Z digest=sha256:633248358f8abf75abd010def1bdaeb4a003d16304498bb77c7a8ad8c326f59a

Observation e137eeb7-5ca9-4f40-9ce0-8a2383bd22c3 · outbound

This paper cites High- fidelity audio compression with improved rvqgan,.

Rate-Aware Learned Speech Compression High- fidelity audio compression with improved rvqgan,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.600706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.220758Z digest=sha256:775bb0e61be9ca0cf827db79246572788a9a8688648e25109ea6d3e370ba894d

Observation acabeb25-fc9b-4d2f-b41f-4d4812811a69 · outbound

This paper cites SEANet: A Multi-modal Speech Enhancement Network.

Rate-Aware Learned Speech Compression SEANet: A Multi-modal Speech Enhancement Network

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.226962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.226962Z digest=sha256:b00947a5636cb89a65fc6c9f05a7dfa0ef5cb4fad7ae9fd85227e0f82ed7ef2e

Observation ed381c3b-3dea-4cee-84f1-f99d7a5380d3 · outbound

This paper cites Real-time speech frequency bandwidth extension,.

Rate-Aware Learned Speech Compression Real-time speech frequency bandwidth extension,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.583163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.232669Z digest=sha256:9f38845f0dccaa834aac8b105bedc2946e4f540ff8b3cbcb10d7223071c2aced

Observation dae36b5f-f77d-4557-b950-f751367429b3 · outbound

This paper cites Long short-term memory,.

Rate-Aware Learned Speech Compression Long short-term memory,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.237799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.237799Z digest=sha256:e731e57bc46f4cd4be08fc982502ae8b78defe7a1aff06a50f13d1451e94eb55

Observation 6029b862-5fcf-4332-a490-2d5cd8c63ec9 · outbound

This paper cites Attention is all you need,.

Rate-Aware Learned Speech Compression Attention is all you need,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.242381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.242381Z digest=sha256:09c4dd1b3493f3e46afc1425456dc3c612e842e9de002035abd865a6f6c11c43

Observation 9b8a1c25-e5cd-40e5-931e-35e85fad88e3 · outbound

This paper cites Melgan: Generative adversarial networks for conditional waveform synthesis,.

Rate-Aware Learned Speech Compression Melgan: Generative adversarial networks for conditional waveform synthesis,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.246983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.246983Z digest=sha256:a5b02c3e2a51bff07f2ec3bf26f650c06285bca6d7d4935c5a187b2674561573

Observation e115d7ab-84cb-498f-9c20-8672a0849123 · outbound

This paper cites Language-Codec: Bridging Discrete Codec Representations and Speech Language Models.

Rate-Aware Learned Speech Compression Language-Codec: Bridging Discrete Codec Representations and Speech Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.251505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.251505Z digest=sha256:aa55823145e3b9866ad7553b31bddf96c07f5d5ab64efd6395d5a40b113561f0

Observation b38c3379-53e0-4b0d-b6a9-006a91d5b608 · outbound

This paper cites Variational image compression with a scale hyperprior.

Rate-Aware Learned Speech Compression Variational image compression with a scale hyperprior

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.256882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.256882Z digest=sha256:ce0a87f529f80fdf0b4b2c93b99b7ae3bfc8d66b1e68a52141f512c52256bc06

Observation 4c81a297-3db8-4b64-9aab-33a8a9bf8ab2 · outbound

This paper cites Learned image com- pression with discretized gaussian mixture likelihoods and attention modules,.

Rate-Aware Learned Speech Compression Learned image com- pression with discretized gaussian mixture likelihoods and attention modules,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.261640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.261640Z digest=sha256:eecf3b542463a29e4e1fc7b62a5428327d62606ba936fc1c7cc3e6363f6aa658

Observation 15aa9973-15f9-4e10-9b7f-5fb4f07ffb7f · outbound

This paper cites Channel-wise autoregressive entropy models for learned image compression,.

Rate-Aware Learned Speech Compression Channel-wise autoregressive entropy models for learned image compression,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.520646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.265913Z digest=sha256:7b4b30aa65879b014ed6c548820705d44a97497bed345049bce34031dc863159

Observation 17aeacd7-2aa7-4b44-8e6f-760056901a2e · outbound

This paper cites Elic: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,.

Rate-Aware Learned Speech Compression Elic: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.270460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.270460Z digest=sha256:4d1c5f34d7cf92a90018a482839ec937d598c4b6be895dcd1136f5352b4b9a1d

Observation a562ecd5-f337-48a3-bda4-a0307e6c41a9 · outbound

This paper cites Learned image compression with mixed transformer-cnn architectures,.

Rate-Aware Learned Speech Compression Learned image compression with mixed transformer-cnn architectures,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.492017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.274933Z digest=sha256:503e15099081b07a2f67e7f64975a2972186720d807350616900fb6e2eb0c0ef

Observation 50d5abd5-c9c3-4594-90fa-9077d99611c3 · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

Rate-Aware Learned Speech Compression RWKV: Reinventing RNNs for the Transformer Era

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.279555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.279555Z digest=sha256:f8b06488e6345fcacd07b5916c8d8611815282b2e886148f30293c884e514659

Observation 79188123-8e45-455b-9f62-fffd7fb1c869 · outbound

This paper cites Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence.

Rate-Aware Learned Speech Compression Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.284388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.284388Z digest=sha256:c9f65442f27640907f779a5a6a4f1fd283d541f48068bda1f6536bafc12b6611

Observation ee4ea5b8-3ebb-472d-bfa1-78ecf644c5bf · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

Rate-Aware Learned Speech Compression LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.289645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.289645Z digest=sha256:fc3ce048cd423f50cfe0a3f4b5d30329f9c0cf215ff8654689bad59fb82a0a4b

Observation f27bea2f-cf59-45fa-994a-c4ae2bee2ba9 · outbound

This paper cites Visqol: an objective speech quality model,.

Rate-Aware Learned Speech Compression Visqol: an objective speech quality model,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:41:19.472584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T17:41:19.294217Z digest=sha256:f3a65d91aef02315d28f2171c1caadbdf8d125af78bf0e3543bffbceb70e1dc3

Observation 02a8ef96-e531-4dd3-87d7-6e3d496dda9f · outbound

This paper cites Per- ceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,.

Rate-Aware Learned Speech Compression Per- ceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T17:41:19.298523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:41:19.298523Z digest=sha256:9d780301f0b71b6aed852c5fcf856afbf5f11a5d8c02c803c0ed3d981e1d45ee

Pith citing papers

No inbound Pith citation observations are available.