Pith. sign in

Paper Citation Record · LEDGER

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

As of 9 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 4 inbound Pith citation observations for arXiv:2505.19774.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19774 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:49.357205Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:46.355892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:28:39.508010Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved21
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f25726a-38b9-46a3-aa95-c855e7e20be6 · outbound

This paper cites Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:53.104993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.192370Z digest=sha256:7547f96f4c95aaac54a29aeab7eb33af886de7cf42b7629a1d8305b81492b4d9

Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · outbound

This paper cites DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.355892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.355892Z digest=sha256:a53013496a6a9893486935452fde6dfc632b6c9e72525180d478e8d240666c7d

Observation a6ad7272-62b5-4be4-9a9a-02d1a43c21e6 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:52.685483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.467017Z digest=sha256:2718806fc21ff19fffaf53f6af097bc67cb3afe105513c92cca533df4f761125

Observation fa9dcbb8-7061-43e1-b826-bdaec77172bb · outbound

This paper cites Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:12:52.286979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.646277Z digest=sha256:fb94663260d36b0feeed6c97d3ff74d73344e4c665c06bacb34d930af402fd8b

Observation 46293429-1243-40a3-9b44-805d29587b3a · outbound

This paper cites Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:52.496675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.552005Z digest=sha256:fb3a5faacca670c00640dc818c1abfe837f4b532bef16c55f176b285e18578fe

Observation 2f177c02-655e-4476-ab67-569e5eed50a5 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:52.908471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.274896Z digest=sha256:6d8741fe7770eaa8e387fe8d9f86fd5ccd0d946fc4cd15277808de421e3808b6

Observation c65c6e3a-f562-467d-81ce-1e2a83a8e978 · outbound

This paper cites Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4).

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:52.127133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.744428Z digest=sha256:5df61c9d02e0ce201fd4c798c7e3b15453711472df24eaba20678713d9348c45

Observation ff259256-d9db-4d11-929e-1baa37b58f5e · outbound

This paper cites In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode

Reference 8

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:12:51.897019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.835187Z digest=sha256:1b5b4c443dc1c8fa193ee4cabe92d381ea49bcf779431a51a2ef10ee433c484d

Observation 3df5a67f-bfa2-4c30-b26b-c66d3e4dc616 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:51.660143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:46.895063Z digest=sha256:688ac26ec7f2353c0e990d567b6007b324ab5ddfcbe2c6a43223963e19f900dd

Observation c9700196-b019-4ea0-b376-8035a5b3ad6b · outbound

This paper cites Google usm: Scaling automatic speech recognition beyond 100 languages,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google usm: Scaling automatic speech recognition beyond 100 languages,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.963664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.963664Z digest=sha256:df4e5883b45fcb86ddc6c6d1616d27551c0fa239ae792d2ed62ff8b785b8f7fc

Observation ec2afa29-f28f-4268-8c15-77ae428a0d6f · outbound

This paper cites Self-supervised speech representation learning: A review,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised speech representation learning: A review,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.108528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.108528Z digest=sha256:4e66180e874bbac485309af3d79f75cb77663cc05275d63a6878445815e7e1d7

Observation ac4fef8f-9040-40f2-8b58-b96e78ba18bf · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Robust Speech Recognition via Large-Scale Weak Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.158251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.158251Z digest=sha256:be6f81fafe301c9423131717e5b982ab214d246af5bd19fb9d5c291cc582ed44

Observation 7b8ab252-ca3f-4248-ad46-fc0a3ba2ac4a · outbound

This paper cites Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.243261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.243261Z digest=sha256:6197e48cd4ed10e04ae55c8329bd4852d0f23296d7857642c8e6048a1d92e542

Observation 81e526fa-4163-487f-9b74-bf211c4000bb · outbound

This paper cites Conformer with dual-mode chunked attention for joint online and offline ASR.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer with dual-mode chunked attention for joint online and offline ASR

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:50.153372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:47.355685Z digest=sha256:2a1cbc6cef887b98a76069f61a81e1877244d63d59f13bd59835b0df5f94625a

Observation 4aebe4a4-8b17-490a-ae73-0d5d352b39a3 · outbound

This paper cites Multi- mode transformer transducer with stochastic future context,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi- mode transformer transducer with stochastic future context,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.451479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:47.466237Z digest=sha256:3c84433441e3e9840b162bb9f7ba97a5f5f6812ba7ecaa15d6b32859ca44d5b7

Observation 5b58ca95-c45e-4e83-8170-4a4517a1083b · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.072372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:48.584229Z digest=sha256:db9f5e07d06c4833b54e5043db6f58d91bc69cccb94ef6ca5d6c59aa8377a984

Observation 9acf3d35-c4ac-457a-8387-83306a72b0c6 · outbound

This paper cites Variable Attention Masking for Configurable Transformer Transducer Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Variable Attention Masking for Configurable Transformer Transducer Speech Recognition

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:49.920805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:47.679297Z digest=sha256:ddb77a505650b5d4517e81c09ba5106b471eba5861bd0f7aa798e9ed23ef9c06

Observation 862a48ce-cd21-4cf8-9386-2c06d30700c2 · outbound

This paper cites Sequence transduction with recurrent neural networks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Sequence transduction with recurrent neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.319686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:47.760230Z digest=sha256:b5e5ad1e2b45e1fc12e7dbdb0aed59cde6f988c196de8e57b804462d8055cce5

Observation c822dc6b-1ea6-4b31-a86a-eea374679c94 · outbound

This paper cites SUPERB: Speech processing Universal PERformance Benchmark.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation SUPERB: Speech processing Universal PERformance Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.829556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.829556Z digest=sha256:445e5030257c4f0e9931aa712ef2deca8096113d76aced1a62edbfb11efbc84c

Observation 534ac630-f760-4521-9d3a-06d5baefddae · outbound

This paper cites Self-supervised Learning with Random-projection Quantizer for Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised Learning with Random-projection Quantizer for Speech Recognition

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:49.761320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:47.911681Z digest=sha256:d2f40f7c2e32dac9a71030a30c51a2b1c3b73a418b085d0cdf7758be535a7484

Observation 157c9368-a37c-426c-ae3f-e7a5d50f0181 · outbound

This paper cites Universal-1: Robust and accurate multi- lingual speech-to-text,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Universal-1: Robust and accurate multi- lingual speech-to-text,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.197138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:48.032670Z digest=sha256:e3b17cd1d490bcb5ef7f5bfcff180081cee8e0f18bc3282931f61c9512907202

Observation b7da4468-3076-44dc-a9eb-cc06b1140fcc · outbound

This paper cites ML-SUPERB: Multilingual Speech Universal PERformance Benchmark.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation ML-SUPERB: Multilingual Speech Universal PERformance Benchmark

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.072445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.072445Z digest=sha256:b12c52e4a39e8236285492c8afcc3ba0be3fb649c382f42bc110746d5633cfa8

Observation a09b4471-d3be-4e13-9274-7b338103663e · outbound

This paper cites Lib- rispeech: An asr corpus based on public domain audio books,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Lib- rispeech: An asr corpus based on public domain audio books,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.177512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.177512Z digest=sha256:1f5a7839450ee6bcb222f27960dbd504b4ddd90fc25ac2e6ec23f8bb407596db

Observation 7cca1858-305a-43de-a932-0019aa2902fb · outbound

This paper cites Mls: A large-scale multilingual dataset for speech research,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Mls: A large-scale multilingual dataset for speech research,

Reference 24

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:12:48.292894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.292894Z digest=sha256:87aad9a87398395d96d3b1d95b188761f628681269b7b4db548cd884862814b9

Observation 8907619b-305c-480c-bbdf-d9ac1a0a2de7 · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Common Voice: A Massively-Multilingual Speech Corpus

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.384304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.384304Z digest=sha256:2a7d748448e85c13e197d4fa98a1a02060044721928fc663c747583d14004acf

Observation 52b343b8-f89f-43d4-9dfa-e35a1fa2c00c · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.463964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.463964Z digest=sha256:d7c966bea91c2b17751ba37bcfdf41809640dd91c1bec75e42449ffe50a97469

Observation 02ec53c7-2e9d-4eb7-ba40-45b105b4b1a3 · outbound

This paper cites On the utility of self-supervised mod- els for prosody-related tasks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation On the utility of self-supervised mod- els for prosody-related tasks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.352190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:49.357205Z digest=sha256:af04671ba9bc58b898dfb494215c6855898cf687343385766ad1f2d711d01813

Observation fc9a8ab5-343b-4ba1-ae6c-dedcb0b7b62b · outbound

This paper cites VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.668479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.668479Z digest=sha256:d2f63a7be88f3994fed910d8f9d36c245cd1b13d4ebf1805b3818c614978e9b0

Observation 8c545bf0-d54b-4e6d-a7c6-dcd66f3bcc2b · outbound

This paper cites free public domain audiobooks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation free public domain audiobooks,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.931336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:48.737679Z digest=sha256:68eadb762bb4171f28c2350b2495fd3c58b3d9c459b026228c6eeaa774ddb0cd

Observation 3dc4be5b-c3cf-4eb8-bb71-2b8f6cede795 · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented transformer for speech recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.785451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.785451Z digest=sha256:dedcc45a5fb43ba70aa1629d8128bbf2b43bcfe92077abb9598757137a05093f

Observation 91fb33ab-6728-4d4c-9828-3681a6d691d6 · outbound

This paper cites Long short-term memory,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Long short-term memory,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.979925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.979925Z digest=sha256:2c82c4a8283450e28d941bcecc295d569e38a75db3d8a8deb423b01825e4457a

Observation 4abe35ea-4d8e-4cca-845f-33de1adffd47 · outbound

This paper cites Specaugment: A simple data augmentation method for automatic speech recognition,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Specaugment: A simple data augmentation method for automatic speech recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.723614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:49.032602Z digest=sha256:daeb4e04a4b2387d54ca530d6d87dfd42b5d9769172bbc0b93a9c023b915584c

Observation 4690a66b-005f-4d52-9384-106eafc4bcdc · outbound

This paper cites Msp-podcast corpus: A large naturalistic speech emotional dataset,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Msp-podcast corpus: A large naturalistic speech emotional dataset,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.534911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:49.107373Z digest=sha256:0d769faa7b7659de2f3a54e553bd111d1c5439e610ec21e415a6d8b929e7735b

Observation 98f3f17c-256a-433b-8ca4-77f34dc87fbf · outbound

This paper cites wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.162149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.162149Z digest=sha256:252160e0aaf81e4cd782ce5c86bcc7cfa2d19bc5e53a7797180879b00541a808

Observation 2a313904-5856-45fa-90f8-15d6b4658691 · outbound

This paper cites mHuBERT-147: A Compact Multilingual HuBERT Model.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation mHuBERT-147: A Compact Multilingual HuBERT Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.231627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.231627Z digest=sha256:41f791dbd4b765f32b15a079436ca44958a6fbfba3abb60d13dfcf1751108678

Observation f2af490b-06d7-4ba1-83c2-e87bfd0522b4 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.296912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.296912Z digest=sha256:9682b2520d1abfa35262a96117692fdc98ec7b3f1049fcb74fd28b56fb0adc26

Observation 1e9dd8c2-89ec-41fb-9003-a684f8c85455 · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented Transformer for Speech Recognition

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.881846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.881846Z digest=sha256:c2e496f05ecc33485b4a51926ae01e4dbe637b2eac1c93e7d488ccfa7dd6d58e

Observation 95b117fd-89c3-4d1b-b4e6-a7450b5d471a · outbound

This paper cites Multi-mode Transformer Transducer with Stochastic Future Context.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi-mode Transformer Transducer with Stochastic Future Context

Reference 2021

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:12:50.045674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:12:47.570835Z digest=sha256:ee4b81bd72308b1fd9854aa02b2777dfba09a8a14ba54f66dd19432feb6fa21d

Observation 1a5185df-5743-4224-b270-77f4de0e01cf · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.058460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.058460Z digest=sha256:4831120a69a717e1a57c29fa6442928d45e1bb1eb9625e81552ce6bb1f0b91e9

Pith citing papers

Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · inbound

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation cites this paper.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.355892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.355892Z digest=sha256:a53013496a6a9893486935452fde6dfc632b6c9e72525180d478e8d240666c7d

Observation 4bbfd3a9-aead-4908-a0a2-b5d7e635a8a1 · inbound

Group Relative Policy Optimization for Speech Recognition cites this paper.

Group Relative Policy Optimization for Speech Recognition DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:21.427323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:21.427323Z digest=sha256:5037a379c9eb0bb49a7c1ed541b5c9f2b816bac5d584be69d516cd5d02315b70

Observation 582576fd-6365-410b-8a1b-82e0ab308188 · inbound

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents cites this paper.

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:28:39.509589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T05:32:27.371597Z digest=sha256:2d04308da555be81699b3341c771886e8b7a7713a8b6cb4cf761d4ab526152ff

Observation 8e31b141-02f7-419c-9087-a3b0b05247c2 · inbound

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents cites this paper.

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:43:51.419791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T05:01:44.202712Z digest=sha256:7f6abd23f71a05af780d2aa27cf0844357ae7cd5fdde82b51c1861917a2c4a21