Pith. sign in

Paper Citation Record · LEDGER

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

As of 14 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2505.21578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21578 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:21.783339Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:18.926273Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:49:22.026276Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86228da0-13b9-456f-bba9-b3845192c19b · outbound

This paper cites open-source.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use open-source

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:24.394179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:18.821110Z digest=sha256:9c9cf81d7dcf0325d30d0d2619b9bec22776748639c3bae75a37c9fe1c6ba14d

Observation 2232aa19-443b-418f-9e68-3cfda600b704 · outbound

This paper cites Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:22.160440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:18.926273Z digest=sha256:9f94ec64685d630d3692c2944ba1652e2e9abebb18f6f36dd4768f6d4a3f295b

Observation 3cb75ae8-7806-45ed-8591-b197b1d11987 · outbound

This paper cites dev-other.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use dev-other

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.889982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:19.100288Z digest=sha256:25a6a7481bbf033256d736cc1fc1aaba80c28980a1db2338a50eb607dd1ae74c

Observation 2f61fc9c-bdac-46c9-bef6-fb1aa54c9974 · outbound

This paper cites an unresolved cited work.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Unresolved cited work

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T13:49:23.743847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:19.188453Z digest=sha256:bea4ce7b6ebb7f1d4407ea4c264090a71c578827838fb43c5791c01027daf17e

Observation 6921c4be-1ce9-4e8e-8e98-955461dd57aa · outbound

This paper cites The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.596584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:19.273591Z digest=sha256:2e0db57e0d84d1184fa9b22172bdb569babca64d4d9f01b25982b63ded9e1da4

Observation 7bb05ae4-89c5-4403-a3fd-02391cb282d9 · outbound

This paper cites Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.936707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:20.155009Z digest=sha256:6b8e111a81b25a8e9857ca41069442e9a96fdf80f1e5627628b45891c68f8b89

Observation 3c5899e0-7220-43a1-9edb-d87e6474afe9 · outbound

This paper cites The design for the wall street journal- based csr corpus,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The design for the wall street journal- based csr corpus,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.282469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:19.415192Z digest=sha256:cfc523d3eb91f3d38007fb24d3ecfd877cea6222c07fc20ec1a2ee14c66b8cb5

Observation 21e858af-ea56-477b-b161-b649dc65e523 · outbound

This paper cites Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.109696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:19.540643Z digest=sha256:f50a59ad20bc43a16ce4cd2518638f6a516fac9bf78d87032d5233d79b7365ff

Observation 2833979d-5341-4859-b931-91574bd846d9 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Lib- rispeech: an asr corpus based on public domain audio books,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:19.722000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:19.722000Z digest=sha256:39beb23e2d06137d02999aa40d681ed38343298579375961ddf3365f46575426

Observation e3b28d05-9402-41ee-a114-f234b230419b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Robust speech recognition via large-scale weak supervision,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:19.868605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:19.868605Z digest=sha256:3d4f657266268da28f1f9b412913b7aa8aff336650cba23d21bb518f92ff2cc4

Observation 27e30a3f-6914-4aef-b1de-f9656e08d8ce · outbound

This paper cites Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.016438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.016438Z digest=sha256:a2e65eded743032909d97388e90726d1ac06fe8f0d7d2726b2db1da88aae88e4

Observation 48f8d9f6-776d-4b89-85bd-820eb4c85620 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.093321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.093321Z digest=sha256:82ffcd1c25e87798310aabf4b9c32f8763b55ea1b70b04a54bcbb71795ac49e1

Observation 06f76e27-7e8b-4b08-bc88-fe1b898c0c92 · outbound

This paper cites Yodas: Youtube-oriented dataset for audio and speech,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Yodas: Youtube-oriented dataset for audio and speech,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.351770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.351770Z digest=sha256:929d602dcd8585f4ad11e3e4183c6d9daa9aacb21aa0cca86116495b8dc6fe84

Observation f10eea87-3652-4171-bb8e-2ae55c4a4af5 · outbound

This paper cites Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.780955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:20.544196Z digest=sha256:9bc453154d1a69485050878afdf2c9b5d45b1b31e115099ac8cdd6101318dee5

Observation 287cf320-8e23-46a3-8c7c-cad2c419ef2f · outbound

This paper cites People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:24.045529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:19.029598Z digest=sha256:a686619cd73b02efc59a6a44cf3c7a52f5e00fe189698b20168f17d41f7655f0

Observation d94107cd-9278-4f9c-abfd-72870ccb72d7 · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.652179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.652179Z digest=sha256:e6ee10c9fc46b59a619ce9c9c1e243f102248ae4de1db60b79aedb4c30c5a25c

Observation fc3ba030-0bf9-4472-86b6-df1c827b9927 · outbound

This paper cites SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.806789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.806789Z digest=sha256:b47cecca51d78a1be1f76d5156e1b2b2c37a14de220fe9137ddff839b59b6851

Observation 4794e4e2-7e1b-4afe-96ca-bf96996520d8 · outbound

This paper cites Reproducing whisper-style training using an open-source toolkit and publicly available data,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Reproducing whisper-style training using an open-source toolkit and publicly available data,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.960789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.960789Z digest=sha256:6f7b25bcea18e4e431c142ac91c18d40f98025354bbe5de1e0b5b37c8b1a5514

Observation 9dc565a7-c1be-46a6-9342-50d71a9f83fe · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Common Voice: A Massively-Multilingual Speech Corpus

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.198546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.198546Z digest=sha256:0a8513d6ab6a4383cff2cfb42ae3fd4152046529f30cc864b5f984abeed0f9ca

Observation 39e6ca45-5e08-4ef6-acb3-1f1219484f0a · outbound

This paper cites Open-source conversational ai with speechbrain 1.0,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Open-source conversational ai with speechbrain 1.0,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.623249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:21.352107Z digest=sha256:705e837e1ed8b141c21ae7bc4399b09c9b794c38c621cb1558879df3bb49861e

Observation 80bb602e-4d01-4b5d-82c3-8a56169d6738 · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use MLS: A Large-Scale Multilingual Dataset for Speech Research

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.471158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.471158Z digest=sha256:7d3a45ccfbed9b39728092b02c5d6f1915efe69ac155efe3774ed2d2d3dfe338

Observation 44b33b1d-a72a-4c23-94b9-0f3acd44b9ea · outbound

This paper cites V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.464684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:21.645978Z digest=sha256:c72b15157397c8a25401b1885b61596005423d3a03c6de78fd85b35ec66a6b1d

Observation a99041cf-c46f-4689-bd42-82de3fd80fbb · outbound

This paper cites Joint ctc-attention based end-to-end speech recognition using multi-task learning,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Joint ctc-attention based end-to-end speech recognition using multi-task learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.319240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:21.783339Z digest=sha256:026fea572fefcb7d608e03d9ead35bb719f98b675455eda96ffd285453f69867

Pith citing papers

Observation 2232aa19-443b-418f-9e68-3cfda600b704 · inbound

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use cites this paper.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:22.160440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T13:49:18.926273Z digest=sha256:9f94ec64685d630d3692c2944ba1652e2e9abebb18f6f36dd4768f6d4a3f295b