Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:21.783339Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2505.21578.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:21.783339Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:18.926273Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T13:49:22.026276Z
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 86228da0-13b9-456f-bba9-b3845192c19b · outbound
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2232aa19-443b-418f-9e68-3cfda600b704 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3cb75ae8-7806-45ed-8591-b197b1d11987 · outbound
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f61fc9c-bdac-46c9-bef6-fb1aa54c9974 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6921c4be-1ce9-4e8e-8e98-955461dd57aa · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7bb05ae4-89c5-4403-a3fd-02391cb282d9 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c5899e0-7220-43a1-9edb-d87e6474afe9 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The design for the wall street journal- based csr corpus,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 21e858af-ea56-477b-b161-b649dc65e523 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2833979d-5341-4859-b931-91574bd846d9 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Lib- rispeech: an asr corpus based on public domain audio books,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3b28d05-9402-41ee-a114-f234b230419b · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Robust speech recognition via large-scale weak supervision,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27e30a3f-6914-4aef-b1de-f9656e08d8ce · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48f8d9f6-776d-4b89-85bd-820eb4c85620 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06f76e27-7e8b-4b08-bc88-fe1b898c0c92 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Yodas: Youtube-oriented dataset for audio and speech,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f10eea87-3652-4171-bb8e-2ae55c4a4af5 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 287cf320-8e23-46a3-8c7c-cad2c419ef2f · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d94107cd-9278-4f9c-abfd-72870ccb72d7 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc3ba030-0bf9-4472-86b6-df1c827b9927 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4794e4e2-7e1b-4afe-96ca-bf96996520d8 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Reproducing whisper-style training using an open-source toolkit and publicly available data,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dc565a7-c1be-46a6-9342-50d71a9f83fe · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Common Voice: A Massively-Multilingual Speech Corpus
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39e6ca45-5e08-4ef6-acb3-1f1219484f0a · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Open-source conversational ai with speechbrain 1.0,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80bb602e-4d01-4b5d-82c3-8a56169d6738 · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use MLS: A Large-Scale Multilingual Dataset for Speech Research
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44b33b1d-a72a-4c23-94b9-0f3acd44b9ea · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a99041cf-c46f-4689-bd42-82de3fd80fbb · outbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Joint ctc-attention based end-to-end speech recognition using multi-task learning,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2232aa19-443b-418f-9e68-3cfda600b704 · inbound
Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.