Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:49.357205Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 4 inbound Pith citation observations for arXiv:2505.19774.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:49.357205Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:46.355892Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T16:28:39.508010Z
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5f25726a-38b9-46a3-aa95-c855e7e20be6 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6ad7272-62b5-4be4-9a9a-02d1a43c21e6 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa9dcbb8-7061-43e1-b826-bdaec77172bb · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 46293429-1243-40a3-9b44-805d29587b3a · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f177c02-655e-4476-ab67-569e5eed50a5 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c65c6e3a-f562-467d-81ce-1e2a83a8e978 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4)
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff259256-d9db-4d11-929e-1baa37b58f5e · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3df5a67f-bfa2-4c30-b26b-c66d3e4dc616 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c9700196-b019-4ea0-b376-8035a5b3ad6b · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google usm: Scaling automatic speech recognition beyond 100 languages,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec2afa29-f28f-4268-8c15-77ae428a0d6f · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised speech representation learning: A review,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac4fef8f-9040-40f2-8b58-b96e78ba18bf · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Robust Speech Recognition via Large-Scale Weak Supervision
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b8ab252-ca3f-4248-ad46-fc0a3ba2ac4a · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81e526fa-4163-487f-9b74-bf211c4000bb · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer with dual-mode chunked attention for joint online and offline ASR
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4aebe4a4-8b17-490a-ae73-0d5d352b39a3 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi- mode transformer transducer with stochastic future context,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5b58ca95-c45e-4e83-8170-4a4517a1083b · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Fleurs: Few-shot learning evaluation of universal representations of speech,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9acf3d35-c4ac-457a-8387-83306a72b0c6 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Variable Attention Masking for Configurable Transformer Transducer Speech Recognition
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 862a48ce-cd21-4cf8-9386-2c06d30700c2 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Sequence transduction with recurrent neural networks,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c822dc6b-1ea6-4b31-a86a-eea374679c94 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation SUPERB: Speech processing Universal PERformance Benchmark
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 534ac630-f760-4521-9d3a-06d5baefddae · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised Learning with Random-projection Quantizer for Speech Recognition
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 157c9368-a37c-426c-ae3f-e7a5d50f0181 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Universal-1: Robust and accurate multi- lingual speech-to-text,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7da4468-3076-44dc-a9eb-cc06b1140fcc · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation ML-SUPERB: Multilingual Speech Universal PERformance Benchmark
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a09b4471-d3be-4e13-9274-7b338103663e · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Lib- rispeech: An asr corpus based on public domain audio books,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cca1858-305a-43de-a932-0019aa2902fb · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Mls: A large-scale multilingual dataset for speech research,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8907619b-305c-480c-bbdf-d9ac1a0a2de7 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Common Voice: A Massively-Multilingual Speech Corpus
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52b343b8-f89f-43d4-9dfa-e35a1fa2c00c · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02ec53c7-2e9d-4eb7-ba40-45b105b4b1a3 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation On the utility of self-supervised mod- els for prosody-related tasks,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fc9a8ab5-343b-4ba1-ae6c-dedcb0b7b62b · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c545bf0-d54b-4e6d-a7c6-dcd66f3bcc2b · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation free public domain audiobooks,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3dc4be5b-c3cf-4eb8-bb71-2b8f6cede795 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented transformer for speech recognition,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91fb33ab-6728-4d4c-9828-3681a6d691d6 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Long short-term memory,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4abe35ea-4d8e-4cca-845f-33de1adffd47 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Specaugment: A simple data augmentation method for automatic speech recognition,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4690a66b-005f-4d52-9384-106eafc4bcdc · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Msp-podcast corpus: A large naturalistic speech emotional dataset,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 98f3f17c-256a-433b-8ca4-77f34dc87fbf · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a313904-5856-45fa-90f8-15d6b4658691 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation mHuBERT-147: A Compact Multilingual HuBERT Model
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2af490b-06d7-4ba1-83c2-e87bfd0522b4 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Wavlm: Large-scale self-supervised pre-training for full stack speech processing,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e9dd8c2-89ec-41fb-9003-a684f8c85455 · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented Transformer for Speech Recognition
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95b117fd-89c3-4d1b-b4e6-a7450b5d471a · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi-mode Transformer Transducer with Stochastic Future Context
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1a5185df-5743-4224-b270-77f4de0e01cf · outbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · inbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bbfd3a9-aead-4908-a0a2-b5d7e635a8a1 · inbound
Group Relative Policy Optimization for Speech Recognition DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 582576fd-6365-410b-8a1b-82e0ab308188 · inbound
Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e31b141-02f7-419c-9087-a3b0b05247c2 · inbound
Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.