Pith. sign in

Paper Citation Record · LEDGER

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

As of 8 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2505.21578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21578 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:21.783339Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:18.926273Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:49:22.026276Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86228da0-13b9-456f-bba9-b3845192c19b · outbound

This paper cites open-source.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use open-source

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:24.394179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:18.821110Z digest=sha256:0941eae9dcfa79d2b383fefe05d38d92ea06f21d890542ec6bab98172657b0ff

Observation 2232aa19-443b-418f-9e68-3cfda600b704 · outbound

This paper cites Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:22.160440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:18.926273Z digest=sha256:6bf7cf9d8a08ad8e7514efe4c378f2439fed62faaf9f9958c3b0d2b318700ca5

Observation 3cb75ae8-7806-45ed-8591-b197b1d11987 · outbound

This paper cites dev-other.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use dev-other

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.889982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:19.100288Z digest=sha256:28e1e67f776d29aba3e19ca73d8920236ba7bcc9bcd6ee8cd7a886ad06fc7d1e

Observation 2f61fc9c-bdac-46c9-bef6-fb1aa54c9974 · outbound

This paper cites an unresolved cited work.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Unresolved cited work

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T13:49:23.743847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:19.188453Z digest=sha256:713baf22d519ed6d4a17b5215f8bed111d0a49198a55dc25695d0d73d89ee8b2

Observation 6921c4be-1ce9-4e8e-8e98-955461dd57aa · outbound

This paper cites The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.596584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:19.273591Z digest=sha256:d0f253ce7ac0896470b74e6e118adc1fd953d149c706080dd175f0af0206253d

Observation 7bb05ae4-89c5-4403-a3fd-02391cb282d9 · outbound

This paper cites Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.936707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:20.155009Z digest=sha256:f04c63113f51e61fed28a200ccdc84ab8b1b9107d1cef6cb9aab285424866049

Observation 3c5899e0-7220-43a1-9edb-d87e6474afe9 · outbound

This paper cites The design for the wall street journal- based csr corpus,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The design for the wall street journal- based csr corpus,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.282469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:19.415192Z digest=sha256:6095727242d227545812aa0c50b06aeb0874e8a52145fdbf340dafccd97c4a14

Observation 21e858af-ea56-477b-b161-b649dc65e523 · outbound

This paper cites Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.109696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:19.540643Z digest=sha256:e61b24df2e195ebddb49cf9fe53694bd78a379d7729bdcb9ac2f128659d27729

Observation 2833979d-5341-4859-b931-91574bd846d9 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Lib- rispeech: an asr corpus based on public domain audio books,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:19.722000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:19.722000Z digest=sha256:b5941458e8b445c944992f08bc3b8cf7c9ea2c1135f4074df710a44cfb8ed2c0

Observation e3b28d05-9402-41ee-a114-f234b230419b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Robust speech recognition via large-scale weak supervision,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:19.868605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:19.868605Z digest=sha256:4d7a8b7fab7d4de73259a51524f8af68de7895a704166456636732b37dc93b21

Observation 27e30a3f-6914-4aef-b1de-f9656e08d8ce · outbound

This paper cites Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.016438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.016438Z digest=sha256:4940eec437ab38480bf447c2d2bf4fd9aca2b0b52f9697acc83613889d3dd465

Observation 48f8d9f6-776d-4b89-85bd-820eb4c85620 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.093321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.093321Z digest=sha256:6c44744c3f5b8df8991079abe87435d9a07eb9231ca2a21adc47d029628bd6c7

Observation 06f76e27-7e8b-4b08-bc88-fe1b898c0c92 · outbound

This paper cites Yodas: Youtube-oriented dataset for audio and speech,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Yodas: Youtube-oriented dataset for audio and speech,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.351770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.351770Z digest=sha256:4c6b6f9b562231e18d25a797a79e969459aa4f6959a594b8fa95d05b16876db1

Observation f10eea87-3652-4171-bb8e-2ae55c4a4af5 · outbound

This paper cites Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.780955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:20.544196Z digest=sha256:ab99b8ffbcc59637d40f476520340703f0ff2a338b6d982bf32a138c358d6f38

Observation 287cf320-8e23-46a3-8c7c-cad2c419ef2f · outbound

This paper cites People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:24.045529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:19.029598Z digest=sha256:561dc6ac67ee29add582577d815561a544b397d66c86844b26d58786620304f8

Observation d94107cd-9278-4f9c-abfd-72870ccb72d7 · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.652179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.652179Z digest=sha256:3161c9f01a3fe0db42f9785f6fefbd4bf4d31757ccf68c58d2507a9a852e8ad8

Observation fc3ba030-0bf9-4472-86b6-df1c827b9927 · outbound

This paper cites SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.806789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.806789Z digest=sha256:f225ee3c5ef9177f976d980b948ab5da806856616d20a8af88dcc4d725c38a4a

Observation 4794e4e2-7e1b-4afe-96ca-bf96996520d8 · outbound

This paper cites Reproducing whisper-style training using an open-source toolkit and publicly available data,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Reproducing whisper-style training using an open-source toolkit and publicly available data,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.960789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.960789Z digest=sha256:1c8644e30d6e460f576236dd36b289091a9e38420558791f0cd5963dd907e369

Observation 9dc565a7-c1be-46a6-9342-50d71a9f83fe · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Common Voice: A Massively-Multilingual Speech Corpus

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.198546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.198546Z digest=sha256:9027def6d4a532e851e7e479657f0d39e1437d8cb70c428c0a2e67bdad2bf633

Observation 39e6ca45-5e08-4ef6-acb3-1f1219484f0a · outbound

This paper cites Open-source conversational ai with speechbrain 1.0,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Open-source conversational ai with speechbrain 1.0,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.623249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:21.352107Z digest=sha256:3516462f1a5fedc0687a1185fc02fca102cdd038eda2267541f2caf75a84d098

Observation 80bb602e-4d01-4b5d-82c3-8a56169d6738 · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use MLS: A Large-Scale Multilingual Dataset for Speech Research

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.471158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.471158Z digest=sha256:d9d2cc3831915a139744454b19ec10299f95131b3f499e37311a1ce1958e0d7e

Observation 44b33b1d-a72a-4c23-94b9-0f3acd44b9ea · outbound

This paper cites V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.464684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:21.645978Z digest=sha256:4aaceae6cae5be5e40c1f200e403a59ecc2f3a7fd85b42a4fdb6badfa5674f9d

Observation a99041cf-c46f-4689-bd42-82de3fd80fbb · outbound

This paper cites Joint ctc-attention based end-to-end speech recognition using multi-task learning,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Joint ctc-attention based end-to-end speech recognition using multi-task learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.319240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:21.783339Z digest=sha256:a6a2b9685f9833532af8fa25ddd12a4ad4c7306259dc4c52af0a853dea1064b3

Pith citing papers

Observation 2232aa19-443b-418f-9e68-3cfda600b704 · inbound

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use cites this paper.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:22.160440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:49:18.926273Z digest=sha256:6bf7cf9d8a08ad8e7514efe4c378f2439fed62faaf9f9958c3b0d2b318700ca5