Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2111.09344.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:13:57.274992Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T20:10:07.966776Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 424a669b-2016-4d62-a083-84d59e362c72 · inbound
WavChat: A Survey of Spoken Dialogue Models The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52b343b8-f89f-43d4-9dfa-e35a1fa2c00c · inbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d94107cd-9278-4f9c-abfd-72870ccb72d7 · inbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ec2465d-b686-4696-a457-74c0a22ec0aa · inbound
OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b443f9f-0199-4440-9394-983dbf2e8ece · inbound
Whale: Large-Scale multilingual ASR model with w2v-BERT and E-Branchformer with large speech data The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 007423dc-82e9-40c5-be88-0607471da412 · inbound
Ming-Omni: A Unified Multimodal Model for Perception and Generation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82dc9500-d820-4672-8a3b-a1b4f3ff12c7 · inbound
Instituto de Telecomunica\c{c}\~oes at IWSLT 2025: Aligning Small-Scale Speech and Language Models for Speech-to-Text Learning The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f9a1363-5163-4ac3-aa36-633587a2a7a7 · inbound
Analyzing and Fine-Tuning Whisper Models for Multilingual Pilot Speech Transcription in the Cockpit The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1116d9ff-1bb6-45fe-8c3b-2d690506a98d · inbound
TTS-1 Technical Report The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3401423f-1608-4c4e-bf02-fad5fe8224fa · inbound
Group Relative Policy Optimization for Speech Recognition The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11f03851-e677-4215-b44e-f17e27cdb100 · inbound
An Empirical Analysis of Discrete Unit Representations in Speech Language Modeling Pre-training The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f3f1d5d-1fc1-4078-aac8-50ad37a2817d · inbound
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f08ae745-8c0c-45c9-8e5d-5c4ce0d8d8ab · inbound
Swivuriso: The South African Next Voices Multilingual Speech Dataset The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7b854e1-cb19-4514-8f8b-5dff47f5ff84 · inbound
Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bb045c2c-b638-4482-a623-086d0037e76f · inbound
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 125
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 068edde5-1384-476f-b6df-1f05db66a483 · inbound
RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation be8a2da7-1b27-480e-87ce-879cc3e59f4b · inbound
RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 49cb625b-4304-4618-9094-15bb842ea5f8 · inbound
RADAR Challenge 2026: Robust Audio Deepfake Recognition under Media Transformations The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 167ab0d4-6cc9-48fb-9d09-f00e78cea41b · inbound
A Semi-Supervised Framework for Speech Confidence Detection using Whisper The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0d7c2b33-d648-40b7-8bcc-e3d509e1b74f · inbound
Raon-OpenTTS: Open Models and Data for Robust Text-to-Speech The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a8d9ac97-39d3-446e-bb53-a2f48a6c3582 · inbound
SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation df4de4f9-255a-4315-90ca-f4423eb35d0b · inbound
Interleaved Speech Language Models Latently Work In Text The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6a8b02e2-f9f1-4c90-8e35-fb7561cdf825 · inbound
From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b47761c7-f254-4d3a-9afa-934c9feee09d · inbound
Listen, Think, Transcribe: Continuous Latent Test-Time Scaling for ASR The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.