Pith. sign in

Paper Citation Record · LEDGER

SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2104.02133.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.02133 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:20.806789Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T14:49:36.377185Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fc3ba030-0bf9-4472-86b6-df1c827b9927 · inbound

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use cites this paper.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.806789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.806789Z digest=sha256:dfc7369268cb8f9b4728850d5966588a75e2f6257c0064c33e0e7d6f77e34a96

Observation 77bf07e5-9d02-4ed0-a257-5038fcbd6f38 · inbound

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning cites this paper.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:19.793022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:19.793022Z digest=sha256:e3187f1b94b8c0ea360c0d1be3920f82eee1930bb7e808d35ce7c7c609d75335

Observation 730e33a4-5b91-4288-80a0-6faee9c852fa · inbound

Analyzing and Fine-Tuning Whisper Models for Multilingual Pilot Speech Transcription in the Cockpit cites this paper.

Analyzing and Fine-Tuning Whisper Models for Multilingual Pilot Speech Transcription in the Cockpit SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:12.145519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:12.145519Z digest=sha256:96803ab90f4bab5372e9cf6a35c8f1c899b0c8efb549e870e73ac8532b8598e4

Observation 58438ae6-60c1-4a4a-816c-8a3a60247d81 · inbound

CAM\~OES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese cites this paper.

CAM\~OES: A Comprehensive Automatic Speech Recognition Benchmark for European Portuguese SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:39:21.241548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:39:21.241548Z digest=sha256:76b2823f5d2cdcbd757784a7c04d76d151bbe1007d0cc19eeba1b830ce479951

Observation 45cc14a7-fb7e-4e29-a232-d72a61ef2a0d · inbound

OLMoASR: Open Models and Data for Training Robust Speech Recognition Models cites this paper.

OLMoASR: Open Models and Data for Training Robust Speech Recognition Models SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:49:36.379813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T14:49:35.843471Z digest=sha256:c31339a4f2812e308244abc5890d9535bf8a01dcea32cb00e8935b305ae6ecd9