Pith. sign in

Paper Citation Record · LEDGER

DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2012.06659.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2012.06659 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:25:27.747596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T02:19:24.235423Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b37905d6-6ee5-4044-bfe4-61ea1b956b0a · inbound

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization cites this paper.

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:52:26.342390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T02:51:33.582296Z digest=sha256:99803db205ad3ba9c7dbc55be3836b2a3b0ac8dd37203fe28f29d88d377a2a39

Observation ebc7b7de-f594-46cd-9f0d-a05eb86275a0 · inbound

Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers cites this paper.

Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:25:27.747596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:25:27.747596Z digest=sha256:a496749d0eae69f2e1e173aa59247a11fa166c56775f8b925403271f1cc95a66

Observation c1407493-1c97-4015-b414-9db29e952f4e · inbound

Pitch Accent Detection improves Pretrained Automatic Speech Recognition cites this paper.

Pitch Accent Detection improves Pretrained Automatic Speech Recognition DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T23:51:12.772243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:51:12.772243Z digest=sha256:0b8da23f286ed545275aff5bf1941a476fee609d58efb265a27cce3fd8e8d495

Observation 4ecb0655-e937-4edd-8b53-e1d3bff6648f · inbound

A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection cites this paper.

A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:26:22.121305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:25:27.137423Z digest=sha256:0f7dd5c0ae05b81ed5c4870c3126eb29fb2ec8fa1b66a0513312b209e30582f4

Observation fad97e3a-982c-4cf2-8af7-08ec291c8d3e · inbound

S-JEPA : Soft Clustering Anchors for Self-Supervised Speech Representation Learning cites this paper.

S-JEPA : Soft Clustering Anchors for Self-Supervised Speech Representation Learning DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:19:24.237120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T19:46:47.653439Z digest=sha256:fb538401ab54a1f9ea75e91b658b536c5690858b3f1d34eec614a0340e1247f0

Observation d71576e8-0795-4c8d-a894-ae4b6cc79834 · inbound

OLIVE: View-Augmented Latent Prediction with Waveform Reconstruction for Speech SSL cites this paper.

OLIVE: View-Augmented Latent Prediction with Waveform Reconstruction for Speech SSL DeCoAR 2.0: Deep Contextualized Acoustic Representations with Vector Quantization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:24:19.499154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T06:17:41.860360Z digest=sha256:69984d4eb4d47622d86822c369d8797ff79f7f6c0c2f86178633f15a98db9bca