Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2101.00390.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:03:32.780424Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:59:42.926035Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b5f067a1-5f74-48e5-bc36-15b8f732aaea · inbound
Gemini: A Family of Highly Capable Multimodal Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 114
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df8b7a56-ea67-465b-8f34-2edfe29b78dc · inbound
XAttnMark: Learning Robust Audio Watermarking with Cross-Attention VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 444d5cec-fedb-4c44-87c4-6ac9730a923b · inbound
Kimi-Audio Technical Report VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cfacb16-af02-45e6-9163-d43e839a6056 · inbound
HPP-Voice: A Large-Scale Evaluation of Speech Embeddings for Multi-Phenotypic Classification VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee506292-b81c-45b4-8312-f35493fea5df · inbound
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12b32fd-de54-4556-817b-b2c42d11dca1 · inbound
From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc9a8ab5-343b-4ba1-ae6c-dedcb0b7b62b · inbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62cf6209-dec3-431b-98bc-12855fdd7f48 · inbound
MFLA: Monotonic Finite Look-ahead Attention for Streaming Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c5bdbc-2ebc-4811-8a2f-a6b72e0b28e1 · inbound
Unified Semi-Supervised Pipeline for Automatic Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d90560f-7ae2-47ac-a5cf-b99fa37eaee7 · inbound
Instituto de Telecomunica\c{c}\~oes at IWSLT 2025: Aligning Small-Scale Speech and Language Models for Speech-to-Text Learning VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e9d6ea3-8a1f-4ed8-8865-22e14f8480c2 · inbound
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ece12cd-30a7-43f6-b941-35aafc10cc84 · inbound
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d9537f6-57df-417c-a1d3-c98e3510fb15 · inbound
On Barriers to Archival Audio Processing VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6404bd70-32fa-466e-b654-406b91370c41 · inbound
An approach to measuring the performance of Automatic Speech Recognition (ASR) models in the context of Large Language Model (LLM) powered applications VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf6f6e49-c096-41c8-86db-03287d382039 · inbound
Group Relative Policy Optimization for Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d292078-0891-4d23-92c2-4d347df25a3b · inbound
SpeechLLM: Unified Speech and Language Model for Enhanced Multi-Task Understanding in Low Resource Settings VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caa0a545-4f09-45b7-a604-0ee6e3a1e41a · inbound
From perception to production: how acoustic invariance facilitates articulatory learning in a self-supervised vocal imitation model VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda344ad-35fd-40b0-8bde-87d3c6790e7e · inbound
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 512e650d-05d0-4566-8218-f216c51451e2 · inbound
ParsVoice: A Large-Scale Multi-Speaker Persian Speech Corpus for Text-to-Speech Synthesis VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f150e90-e3a9-48ec-90b8-7a5747de2ada · inbound
FastSLM: Hierarchical Temporal Abstraction for Efficient Long-Form Speech Adaptation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21093ffe-41f4-445d-8d76-c11f59339e69 · inbound
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa3ec273-1c9f-4c28-a4d4-a5f4b4cada9b · inbound
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1ff7800-3715-43f2-8791-515dfcb4e63f · inbound
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8f15238-338e-4b4a-bcd5-70bc8f4a5c2d · inbound
Raon-OpenTTS: Open Models and Data for Robust Text-to-Speech VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation afe61cb0-722c-4aac-b777-01fe72d86fb1 · inbound
A Unified and Reproducible Experimentation Framework for Speech Understanding VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50d2ac01-4d92-4ee5-8c71-936d4a543597 · inbound
NaturalFlow: Reducing Disruptive Pauses for Natural Speech Flow in Simultaneous Speech-to-Speech Translation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 796dd2b9-b815-40d1-9b73-63d8316d7d75 · inbound
Interleaved Speech Language Models Latently Work In Text VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2611f189-4158-47dd-b79a-961d6fd9cc4a · inbound
SimulS2ST-Omni: Data-Efficient Streaming Speech-to-Speech Translation via Explicit Trajectory Supervision VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 169
Source-reported events for the cited work
Unavailable: canonical work link unavailable.