Pith. sign in

Paper Citation Record · LEDGER

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning

As of 10 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2506.00338.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00338 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:12:21.174309Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:12:17.177212Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:12:21.729943Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved10
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da98eed8-8609-4fd1-87f4-9a9d2c0507e2 · outbound

This paper cites OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning

Reference 1

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T12:12:21.828937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.177212Z digest=sha256:21cd89f48438053b72cdc8c3e2ee824fe0553fe95d987503f78471720eb054fe

Observation 8fecf34e-4ab5-4486-8dbe-b28d4b53b12e · outbound

This paper cites YODAS data cleaning The raw YODAS data has not undergone a rigorous cleaning process and may contain annotation errors [24].

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning YODAS data cleaning The raw YODAS data has not undergone a rigorous cleaning process and may contain annotation errors [24]

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:12:27.493239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.271724Z digest=sha256:14ac26d8b30f0894c634ba04d543ed283187e7b60b3b4af98a6be0704b31ceae

Observation 62ffcb80-8de8-42e8-be83-0757cdfc274f · outbound

This paper cites transcription.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning transcription

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:12:21.617665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.393419Z digest=sha256:ff1ae9e674348cda6c3ef963ee0f9f0715cbfe307a0a8fe4e3a31e989b62b926

Observation e9ce723c-d9be-4b7c-b87b-6e61c7bcaf36 · outbound

This paper cites We reveal that large-scale web-crawled data contains incorrect lan- guage labels and audio-text misalignments.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning We reveal that large-scale web-crawled data contains incorrect lan- guage labels and audio-text misalignments

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:27.321548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.504498Z digest=sha256:1200e30991e94846bb59d4b85819062f847eedad526976e7ac3c366e91371690

Observation fc45377e-b74f-4575-826c-4b2c2bab28f7 · outbound

This paper cites an unresolved cited work.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:12:27.141555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.586769Z digest=sha256:8380b1e7802eeb32f3b1b5e3f819321b6ad6649537342f2b90f81a9cc74a7236

Observation f81a6e80-6bf0-4300-ad75-9e2560fae4d9 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Robust speech recognition via large-scale weak supervision,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:26.954012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.675610Z digest=sha256:4db3bb8eb759934f1f1fa40e7d0ccd7638dfb30926101a90f126c4cf6c4d5e3f

Observation a49829e8-8114-4c2a-b598-09edd3f5f677 · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:17.790407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:17.790407Z digest=sha256:92d5dd4dfd30f5a1618414e08b230cf04983e4230d0007227cee327082c9e6d4

Observation 1c293df4-4f20-4953-b75d-8f948b50f06e · outbound

This paper cites Scaling speech technology to 1,000+ languages,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Scaling speech technology to 1,000+ languages,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:26.785010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.904554Z digest=sha256:ac9445b88e76018704ba1dc1dd25e3a2293b88a2ba7fb577e64fe7df45a81c7e

Observation cfbe3c66-a058-4b60-9191-a84e9b7324c4 · outbound

This paper cites Less is more: Accurate speech recognition & translation without web- scale data,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Less is more: Accurate speech recognition & translation without web- scale data,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:26.633111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.996141Z digest=sha256:005d37b258fa2b73de98e254763f6ef6cfccc45cfbfac34c2a43b9e339a39750

Observation 538be67f-567f-4d89-9806-f9d4c5892108 · outbound

This paper cites Reproducing Whisper-Style Training Using an Open-Source Toolkit and Pub- licly Available Data,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Reproducing Whisper-Style Training Using an Open-Source Toolkit and Pub- licly Available Data,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:26.501683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.091981Z digest=sha256:c53962d2aded48404849ec3d3250bc020fc5a39fdec2c2d955851f3fe630475d

Observation cef51312-f6bb-4e3d-9d3e-b61f35ee3ab5 · outbound

This paper cites ESPnet: End- to-End Speech Processing Toolkit,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning ESPnet: End- to-End Speech Processing Toolkit,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:26.377366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.171426Z digest=sha256:e72f4b4d27484df40dd58326e2b430e239ae85d4798689400bd9ae4ea1df0e1b

Observation 42b5f007-b1a7-4cbb-b9b5-7db73be9266f · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:26.201127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.265022Z digest=sha256:bf34252fee673d9d64efd088239e63ba9973e4cff4aa4d40f5541ff243bcff8a

Observation efcccd6c-06d7-4d53-b4eb-9218918c66d4 · outbound

This paper cites Branchformer: Parallel MLP-attention architectures to capture local and global context for speech recognition and understanding,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Branchformer: Parallel MLP-attention architectures to capture local and global context for speech recognition and understanding,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.994805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.358460Z digest=sha256:60ec12df696007bbfff975af32067ac8c978c22c99abb6cab2151cb624459d47

Observation d8abf4c3-f8bf-4c5c-b159-f1501bae89d4 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Zipformer: A faster and better encoder for automatic speech recognition,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.745680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.435408Z digest=sha256:e5481581d865762347e94f26961cefa72edab486bcfdbd6522dd83b324cc0164

Observation cb75094c-5bfc-457f-8ba9-344d1fe37679 · outbound

This paper cites Atten- tion is all you need,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Atten- tion is all you need,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.603809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.504149Z digest=sha256:1caf130f75d40e573806089d49e51d784488a4a96bd4bf221043ff2eb2d1e39a

Observation 9d6c92d4-8967-4486-bcf5-337b9e8732a6 · outbound

This paper cites Squeezeformer: An efficient transformer for automatic speech recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Squeezeformer: An efficient transformer for automatic speech recognition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.467750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.599148Z digest=sha256:04507a3d2d583a265ab4b440df7caeed533132315597bf5c182eb53161b55ac2

Observation fc6f5a37-a3b3-41f0-8d97-7f35e083183d · outbound

This paper cites Fast conformer with linearly scalable attention for efficient speech recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Fast conformer with linearly scalable attention for efficient speech recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.321074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.726297Z digest=sha256:72b2ea5126992441e3d7682ce0d649e8c953b852c103d839c16833d25091d31c

Observation 86c09436-cba7-4031-b368-fd2da1233b7f · outbound

This paper cites Sum- maryMixing: A linear-complexity alternative to self-attention for speech recognition and understanding,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Sum- maryMixing: A linear-complexity alternative to self-attention for speech recognition and understanding,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.178642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.814410Z digest=sha256:833635e72216560fc04b23f9546044cf2937b9553a21f140bf50208d72d748e1

Observation 21568b9a-d2a0-425e-819e-d272968b80f9 · outbound

This paper cites OWSM v3.1: Bet- ter and faster open whisper-style speech models based on E- Branchformer,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning OWSM v3.1: Bet- ter and faster open whisper-style speech models based on E- Branchformer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:25.047851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.852042Z digest=sha256:44f687a7f205fcd12c2ab4fba9299f82a9ed9e8a01d3bde955aa7b12ebd41911

Observation 837310f8-d203-4448-84fb-80b4cf0d328f · outbound

This paper cites E-Branchformer: Branch- former with enhanced merging for speech recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning E-Branchformer: Branch- former with enhanced merging for speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.905555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.855538Z digest=sha256:16992b9c59661658f9506e942ac0d8a55bed7f3039606159a0564257072b5e3a

Observation bfcef2ea-205f-4f2c-abe1-7026a12b9920 · outbound

This paper cites A Comparative Study on E-Branchformer vs Conformer in Speech Recognition, Transla- tion, and Understanding Tasks,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning A Comparative Study on E-Branchformer vs Conformer in Speech Recognition, Transla- tion, and Understanding Tasks,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.764470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.888232Z digest=sha256:881f82212a60638124d04b4ee6300e70061550c9ca419efff08981dd5f21f2d3

Observation 8ca5f411-4012-44a5-9995-68ca9fbad818 · outbound

This paper cites OWSM-CTC: An open encoder-only speech foundation model for speech recognition, translation, and language identification,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning OWSM-CTC: An open encoder-only speech foundation model for speech recognition, translation, and language identification,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.649967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:18.949572Z digest=sha256:3a10059cbb809014d0304bf1d400af09ae7a299f841459771011d642d6388fd1

Observation 56d343c9-0199-4c6e-9ab5-09ac9178e1b8 · outbound

This paper cites Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.544340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.036480Z digest=sha256:090fd610661e7b7ace2f2bc665a74aabd203938f056e50b164ac384df6d604df

Observation ea1b3402-bfe7-4ff4-b7a7-7812ed9f1bed · outbound

This paper cites Unsupervised data selection via discrete speech representation for ASR,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Unsupervised data selection via discrete speech representation for ASR,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.424537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.139139Z digest=sha256:c5d1e104d194649df212ac6184b2cfab7183f49e3612716c59196c543838890b

Observation d971e49d-8a2c-4745-bc46-4fcaf0b7582d · outbound

This paper cites Unsupervised data selec- tion for speech recognition with contrastive loss ratios,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Unsupervised data selec- tion for speech recognition with contrastive loss ratios,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.308824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.209648Z digest=sha256:dbcbc02020eb598163f1b83ceafd87e2094b7834b2d284e0f2f0864de850ac38

Observation 272d04d1-4869-4c60-b6c8-f692c1539935 · outbound

This paper cites Spgispeech: 5, 000 hours of transcribed financial audio for fully formatted end-to-end speech recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Spgispeech: 5, 000 hours of transcribed financial audio for fully formatted end-to-end speech recognition,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:24.147242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.301270Z digest=sha256:1e89b8a5d0f5d4476b55da71b83bafc64ec643775726c8eaf70a84d54dd7e38a

Observation 2c835fcb-4178-443b-bed0-c9e0bb01c651 · outbound

This paper cites Gigaspeech: An evolv- ing, multi-domain ASR corpus with 10, 000 hours of transcribed audio,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Gigaspeech: An evolv- ing, multi-domain ASR corpus with 10, 000 hours of transcribed audio,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:23.980936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.428013Z digest=sha256:aa18e38e3e2432b8223b66272049154aec2d8b6371322f06ea2f6feac41733f8

Observation 9ec2465d-b686-4696-a457-74c0a22ec0aa · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:19.512688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:19.512688Z digest=sha256:994ca39af57d071f3eaf0c87124101779d7a9e0fe61ef2e9854966fc111b906f

Observation 475da79a-c3f1-4fcf-b4df-06d6e81bd146 · outbound

This paper cites YODAS: Youtube-Oriented Dataset for Audio and Speech,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning YODAS: Youtube-Oriented Dataset for Audio and Speech,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:23.795234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.597084Z digest=sha256:853db296910e585ad73b376ba4b1c2495b2fe36694f440f2665288d30bce118e

Observation 13297370-5350-435d-8613-107046f5c0d7 · outbound

This paper cites On the effects of het- erogeneous data sources on speech-to-text foundation models,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning On the effects of het- erogeneous data sources on speech-to-text foundation models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:23.611777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.703596Z digest=sha256:ca3574003b850a2c47d9abe0b7c034217d55aa4eceba091f80aa0c1b8cb9fc3f

Observation 77bf07e5-9d02-4ed0-a257-5038fcbd6f38 · outbound

This paper cites SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:19.793022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:19.793022Z digest=sha256:e3187f1b94b8c0ea360c0d1be3920f82eee1930bb7e808d35ce7c7c609d75335

Observation 4575edf5-26aa-40ee-8c78-c593b0634b56 · outbound

This paper cites MSR-86K: An Evolving, Multilingual Corpus with 86,300 Hours of Transcribed Audio for Speech Recognition Research,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning MSR-86K: An Evolving, Multilingual Corpus with 86,300 Hours of Transcribed Audio for Speech Recognition Research,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:23.463609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.884354Z digest=sha256:de503643f125a90e2316f57ce2f6bebc0c0541e3795a6b178c03aa4b66714190

Observation c8bfbcdb-e401-4f64-8bbb-f01fba88ba5a · outbound

This paper cites Libriheavy: A 50,000 hours asr corpus with punctuation casing and context,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Libriheavy: A 50,000 hours asr corpus with punctuation casing and context,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:23.270870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:19.960379Z digest=sha256:b383a9779bb2a533e0fc66e88127fffd3c5ab0e58c7c8302f2dd52d26609d926

Observation 405fdaac-7213-4d65-b898-1e5404abd152 · outbound

This paper cites GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:20.068090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:20.068090Z digest=sha256:9b26e286ccb77d3f3ced11c6732b37596e36656b81708a9383384204b3a3a8b8

Observation b61a362a-029f-4875-a106-8c302d761e8b · outbound

This paper cites MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Founda- tion Model Training on EU Languages,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Founda- tion Model Training on EU Languages,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:23.100262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:20.156850Z digest=sha256:3f50770dd3a055b6e07793946fdb2327d1ed5c3be114dc1bb5c4110dc06cd1b4

Observation 037e1b2f-f68f-4680-9328-3b76ff9c7ab4 · outbound

This paper cites CTC-Segmentation of Large Corpora for German End-to-End Speech Recognition,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning CTC-Segmentation of Large Corpora for German End-to-End Speech Recognition,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:22.956806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:20.255300Z digest=sha256:21de7cd0141fca212763db3cb55ed96b7e366bcf3a75661525ddbd17d464bd07

Observation ff89a244-e405-41d6-a38f-b4a7e9909921 · outbound

This paper cites Bag of Tricks for Efficient Text Classification.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Bag of Tricks for Efficient Text Classification

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:20.365163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:20.365163Z digest=sha256:2070b6d9bf0cbea2b009976d2ef37f7aa3624b1221905d52d40ea511839a3505

Observation 2207002f-9773-4956-b514-6e8c6148b10c · outbound

This paper cites FastText.zip: Compressing text classification models.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning FastText.zip: Compressing text classification models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:20.479161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:20.479161Z digest=sha256:7eb8ee8b44d99c8c7e1889618cb746905c19366e494931c6478a132e19953f4c

Observation a1278efc-5e62-4b40-95e9-1fc712b23bd1 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning SpeechBrain: A General-Purpose Speech Toolkit

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:20.555853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:20.555853Z digest=sha256:b1659bcec5a25b69447caa0141ad3f443512f94b4d37b5953cc8219fbe6ff975

Observation 39e7fa83-9333-474a-b8a6-85f1307ad3fb · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Common Voice: A Massively-Multilingual Speech Corpus

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:20.661052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:20.661052Z digest=sha256:69fb7accbf36037fe2a8f05389ac001bce6a8eea2ca22832835a7df32b0549e9

Observation 2b5be670-959e-4d36-82bc-9f0fba6c716b · outbound

This paper cites Pytorch: An imperative style, high- performance deep learning library,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Pytorch: An imperative style, high- performance deep learning library,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:22.792853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:20.741492Z digest=sha256:554e33fc1c59e4e48d47640f1c7e4f78f6458350fd1075a9db3032cf97fe4f47

Observation 8e7d6a8a-75ae-48f8-ab33-a35fe2579cab · outbound

This paper cites FlashAttention-2: Faster Attention with Better Paral- lelism and Work Partitioning,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning FlashAttention-2: Faster Attention with Better Paral- lelism and Work Partitioning,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:22.648375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:20.824229Z digest=sha256:b5b3c470702ec0795fa35a1781e53542d0d245ddea8448ad8962f1949959e149

Observation 59c95f76-2645-43e3-8463-5e7bbf4706eb · outbound

This paper cites Decoupled weight decay regular- ization,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Decoupled weight decay regular- ization,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:22.469508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:20.908318Z digest=sha256:9717ab4fd90002bdf47527eac66b781f8c4c40540338c7512630bffe372c9753

Observation 9ab61f24-357c-4dfc-915a-db9a72d7dfc7 · outbound

This paper cites CoV oST 2 and Massively Multilingual Speech Translation,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning CoV oST 2 and Massively Multilingual Speech Translation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:22.314269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:21.008812Z digest=sha256:8baf4fd48a65201dd504060fbcf379f0b91a80891cd1ab967095ebb399456f88

Observation 820967f7-1ead-4b55-b5d4-46b2b45fe252 · outbound

This paper cites FLEURS: Few-Shot Learning Evaluation of Universal Representations of Speech,.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning FLEURS: Few-Shot Learning Evaluation of Universal Representations of Speech,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:22.131499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:21.049823Z digest=sha256:4cf8421851e708b58a1ab51d8c03afe19f042280aef6423e1c6d4dd6da135517

Observation d196b4a7-095d-443c-b774-7b1f45f97f0c · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning MLS: A Large-Scale Multilingual Dataset for Speech Research

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:21.094758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:21.094758Z digest=sha256:ca0b91b222d08696ae5f673908fed117aff85d8deb777fb10c658a4034432ff8

Observation 48a6f4de-4db2-4861-8697-d222a52175a2 · outbound

This paper cites Srivastav, S.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Srivastav, S

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:12:21.979286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:21.174309Z digest=sha256:2cf842fe900db9611013159988659de11395a73b989848a66a2f395289147e49

Pith citing papers

Observation da98eed8-8609-4fd1-87f4-9a9d2c0507e2 · inbound

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning cites this paper.

OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning

Reference 1

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T12:12:21.828937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:12:17.177212Z digest=sha256:21cd89f48438053b72cdc8c3e2ee824fe0553fe95d987503f78471720eb054fe