Pith. sign in

Paper Citation Record · LEDGER

vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:1910.05453.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1910.05453 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:42:56.928928Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:27:34.881381Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3d1b2d55-e44f-4080-9d8a-81bdb4d79a59 · inbound

Multimodal Medical Code Tokenizer cites this paper.

Multimodal Medical Code Tokenizer vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T00:42:56.928928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:42:56.928928Z digest=sha256:e4b41207418ead2bad8ab25fbbd89fcb77b4c79420a554bda7b75da803d225b4

Observation f358c909-a9a3-40c7-8130-9d129917742f · inbound

Bitrate-Controlled Diffusion for Disentangling Motion and Content in Video cites this paper.

Bitrate-Controlled Diffusion for Disentangling Motion and Content in Video vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:44:59.929335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:44:59.929335Z digest=sha256:733694173274904bb57e5b0f0b0d420795f1cb59a82657e7c7de4955a25b69a1

Observation 25d505d1-4f49-424e-9ff5-4b9809c63378 · inbound

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs cites this paper.

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:01:24.421342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T12:57:04.450462Z digest=sha256:e9e2e08ad095203af0750b8235ab5df2b7f5791c721a86322282a9e62ed22d8f

Observation cf743998-8796-4234-9694-97d682b945bc · inbound

Incomplete Multi-View Multi-Label Classification via Shared Codebook and Fused-Teacher Self-Distillation cites this paper.

Incomplete Multi-View Multi-Label Classification via Shared Codebook and Fused-Teacher Self-Distillation vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T16:53:00.174441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T16:50:19.940458Z digest=sha256:796d43ca1979b9493bd0da8bc757202e8336722904ff8e284a2267ef3b3f54f8

Observation 4857b5f0-a324-44d1-9439-9d915dabba4f · inbound

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization cites this paper.

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.887715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T12:25:52.847432Z digest=sha256:db30a64d89a7a5f11bd14ba3f9e24e272b6c777f7c95869677e71299f342abaf

Observation a62efb3b-79da-4838-af2e-43f700c455bc · inbound

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization cites this paper.

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:15:07.879196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T23:14:32.494076Z digest=sha256:c4a2bbabdb8750d4a99bb0590c8a5df276ce650a66d3a38c2c5b8bb26ecfd5a3

Observation 2082989e-05b8-4318-81fd-1a4cd4ee0a7a · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 235

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.749867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:62520cb4d3f6f409e988f154bbbea16e47ea8f37bc5ccb0c7b1a52089496cc81

Observation 76f3b6db-b6af-4849-8009-e46213ca13ca · inbound

How Optimality Structures Sparse Dictionaries: A Theory for Understanding SAE Representations cites this paper.

How Optimality Structures Sparse Dictionaries: A Theory for Understanding SAE Representations vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 272

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:46:26.206836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T11:36:45.020411Z digest=sha256:5a97f14b87164d6bc371537d2cfaed1b38f7def0bcb7e034594c4f5855007a48

Observation db3eab54-383c-4847-9ac9-28123ace8316 · inbound

End-to-End Training for Discrete Token LLM based TTS System cites this paper.

End-to-End Training for Discrete Token LLM based TTS System vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T03:27:34.882863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T15:22:06.893507Z digest=sha256:915cc7678ad988fc162542f05147634017b886c5cebd2403c15b664efb14682e