Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:11:57.955428Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 2 inbound Pith citation observations for arXiv:2412.11272.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:11:57.955428Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T22:16:51.917336Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T22:21:53.649506Z
76 of 76 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e1fee2c3-e370-41a2-9439-93222cf0253f · outbound
WhisperFlow: speech foundation models in real time Accessed: 2024-11-3
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dc2052ab-9b74-4902-b738-bf19f9a44b52 · outbound
WhisperFlow: speech foundation models in real time GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17cf4bb-74a5-464f-ba1c-3fa280105781 · outbound
WhisperFlow: speech foundation models in real time Did you hear that? Adversarial Examples Against Automatic Speech Recognition
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30577563-e2bb-4df3-8748-a55b7caaaadf · outbound
WhisperFlow: speech foundation models in real time Apple macbook air tech specs, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7ca7c6ef-9f0b-4613-9aee-f4876768af59 · outbound
WhisperFlow: speech foundation models in real time Neural Machine Translation by Jointly Learning to Align and Translate
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 246f45ca-6e4b-40d1-b74b-08e9cbebf41d · outbound
WhisperFlow: speech foundation models in real time Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 62593c14-0bf6-46d0-b9da-25d63e01403c · outbound
WhisperFlow: speech foundation models in real time A mathematical theory of adaptive control processes
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 090f84b0-c612-4419-84ef-3452f7ae6132 · outbound
WhisperFlow: speech foundation models in real time Speech recognition for clinical documentation from 1990 to 2018: a systematic review
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 494cbbb8-5a55-4348-8ca2-80fa349759ec · outbound
WhisperFlow: speech foundation models in real time Language models are few-shot learners
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 094d6cd5-e6f3-48a0-adb3-52eb0e1797be · outbound
WhisperFlow: speech foundation models in real time Audio adversarial examples: Targeted attacks on speech-to-text
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 725460a3-7847-4591-949e-10a6cb67273c · outbound
WhisperFlow: speech foundation models in real time https://github.com/corsix/amx, 2022
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3877e876-ef45-4874-afa7-83e9d4a5c606 · outbound
WhisperFlow: speech foundation models in real time In 29th USENIX Security Symposium (USENIX Security 20) , pages 2667–2684, 2020
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1e62c57e-2122-4f3f-a686-70b5754b17a6 · outbound
WhisperFlow: speech foundation models in real time Fleurs: Few-shot learning evaluation of universal representations of speech
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50cbdccf-acf2-4b36-9cee-f03a241e06f0 · outbound
WhisperFlow: speech foundation models in real time Automatic recognition of spoken digits
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e23ba16d-0950-4668-a35b-636f32acc0eb · outbound
WhisperFlow: speech foundation models in real time Bert: Pre- training of deep bidirectional transformers for language understanding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2040508e-5056-4dd7-9985-87d990395ede · outbound
WhisperFlow: speech foundation models in real time Speculative decoding for 2x faster whisper inference, 2023
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1ac83dde-2c25-45d3-b37c-06b8261f611e · outbound
WhisperFlow: speech foundation models in real time ggerganov/llama.cpp, 2022
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 26fb2ca9-a17f-4164-a9f1-041b2550ceb3 · outbound
WhisperFlow: speech foundation models in real time ggerganov/whisper.cpp, 2022
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e82e09b5-dc9e-45b2-9b86-ac3f1a5d35df · outbound
WhisperFlow: speech foundation models in real time Sequence Transduction with Recurrent Neural Networks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 468ba433-b7d5-4625-9837-42621a8a68e5 · outbound
WhisperFlow: speech foundation models in real time Conformer: Convolution-augmented Transformer for Speech Recognition
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bef0cb85-d03c-4b74-af88-aaaacca45f7f · outbound
WhisperFlow: speech foundation models in real time MLX: Efficient and flexible machine learning on apple silicon, 2023
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 645054e9-16f2-4b42-9754-79d1b3eac5a9 · outbound
WhisperFlow: speech foundation models in real time Speech understanding systems: Summary of results of the five-year research effort, 1976
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20530a75-8258-4a5c-9c6a-e56172710f77 · outbound
WhisperFlow: speech foundation models in real time Streaming end-to-end speech recognition for mobile devices
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2344ecb9-9d73-415c-a5be-0642d983e51b · outbound
WhisperFlow: speech foundation models in real time Ted-lium 3: Twice as much data and corpus repartition for experiments on speaker adaptation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c88a3b4b-3859-48d0-9a3c-40afca78a3ff · outbound
WhisperFlow: speech foundation models in real time Advances in Joint CTC-Attention based End-to-End Speech Recognition with a Deep CNN Encoder and RNN-LM
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e55ac934-d788-4dcf-988d-6884d4947301 · outbound
WhisperFlow: speech foundation models in real time Swapadvisor: Pushing deep learning beyond the gpu memory limit via smart swapping
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9401c046-5b23-430e-af94-ab6774ffb69f · outbound
WhisperFlow: speech foundation models in real time Deepum: Tensor migration and prefetching in unified memory
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 091c9561-37b9-4a0e-90de-d888f886085b · outbound
WhisperFlow: speech foundation models in real time Speech and language processing, 2000
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2a4a9b6a-28c3-4fe8-bff0-2a03ec80140a · outbound
WhisperFlow: speech foundation models in real time Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af25c6a9-e662-4d0c-b7ab-f308a048cb80 · outbound
WhisperFlow: speech foundation models in real time Scaling Laws for Neural Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8b0faf0-90df-446a-9a85-f6e0d22b16b4 · outbound
WhisperFlow: speech foundation models in real time Convolution-augmented parameter-efficient fine-tuning for speech recognition
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8133820a-9a97-48b7-a7b1-341eb3ebde10 · outbound
WhisperFlow: speech foundation models in real time Speculative Decoding with Big Little Decoder
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cbfb0bb-9383-4450-a179-4372dfc16d32 · outbound
WhisperFlow: speech foundation models in real time Low-latency sequence-to- sequence speech recognition and translation by partial hypothesis selection
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c0e2a732-22d5-4cbc-95cf-eebe0733d498 · outbound
WhisperFlow: speech foundation models in real time Turning Whisper into Real-Time Transcription System
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e04c2db0-c5c3-4d73-96f9-8ca1ec3b96b6 · outbound
WhisperFlow: speech foundation models in real time Streaming automatic speech recog- nition with the transformer model
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 76335910-b08d-4256-9697-772306a91042 · outbound
WhisperFlow: speech foundation models in real time Universal Adversarial Perturbations for Speech Recognition Systems
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00224c54-9fb2-4696-b3fd-d2add7375363 · outbound
WhisperFlow: speech foundation models in real time There is more than one kind of robustness: Fooling Whisper with adversarial examples
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b053425f-17b8-482c-aebc-27494d4481a5 · outbound
WhisperFlow: speech foundation models in real time Train- ing language models to follow instructions with human feedback
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f28c8d69-25ac-4a81-a9ed-dec47100c045 · outbound
WhisperFlow: speech foundation models in real time Lib- rispeech: an asr corpus based on public domain audio books
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6b295caa-a991-45bb-a78a-0597aaed94e0 · outbound
WhisperFlow: speech foundation models in real time Splitwise: Efficient generative llm inference using phase splitting
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 698b7f08-3809-4cb9-aba8-0c5b0a5ee70a · outbound
WhisperFlow: speech foundation models in real time Branchformer: Parallel mlp-attention architectures to capture local and global context for speech recognition and understanding
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b75de65d-c04d-42fb-9686-e1c49c643009 · outbound
WhisperFlow: speech foundation models in real time OWSM-CTC: An Open Encoder-Only Speech Foundation Model for Speech Recognition, Translation, and Language Identification
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 138f80b3-879a-439e-b263-28b3ba8a0e7f · outbound
WhisperFlow: speech foundation models in real time OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98f7d38c-4830-41c1-913c-6c03fff58770 · outbound
WhisperFlow: speech foundation models in real time Reproducing whisper-style training using an open-source toolkit and publicly available data
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d94484bb-c132-4fd9-b332-6f766f875476 · outbound
WhisperFlow: speech foundation models in real time Speech percep- tion at the interface of neurobiology and linguistics
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 251e98f7-4a70-411e-b3c8-3de52bddfbf9 · outbound
WhisperFlow: speech foundation models in real time Improving language understanding by generative pre-training
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cb8518c-9b3a-4ebd-81fb-8dd3c92faf3f · outbound
WhisperFlow: speech foundation models in real time Robust speech recognition via large-scale weak supervision
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af9b2efb-3254-4f00-91fc-7315dbf3a9db · outbound
WhisperFlow: speech foundation models in real time Language models are unsupervised multitask learners
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07330927-76d8-4dae-a04f-5db36bfcee3a · outbound
WhisperFlow: speech foundation models in real time Controlling Whisper: Universal Acoustic Adversarial Attacks to Control Speech Foundation Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 392bc3de-f999-4476-b06d-b3ef08c5a579 · outbound
WhisperFlow: speech foundation models in real time Muting whisper: A universal acoustic adversarial attack on speech foundation models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 006ed930-ce33-42b2-9c7b-ad34ce27e3dc · outbound
WhisperFlow: speech foundation models in real time Paul Robinson, and Bradley S
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c359b220-7008-4c8f-be90-4b7f55c4234c · outbound
WhisperFlow: speech foundation models in real time Exploring architectures, data and units for streaming end-to-end speech recognition with rnn-transducer
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 93db44f8-3349-4a37-86e6-d80025cece7c · outbound
WhisperFlow: speech foundation models in real time ZeRO-Offload: De- mocratizing Billion-Scale model training
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1a6f2e7e-3144-4ebd-98a7-85a1084640b2 · outbound
WhisperFlow: speech foundation models in real time Enhancing the ted-lium corpus with selected data for language modeling and more ted talks
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 34169d80-0d54-47e2-a1af-2b5148244e3a · outbound
WhisperFlow: speech foundation models in real time Adversarial Attacks Against Automatic Speech Recognition Systems via Psychoacoustic Hiding
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eb371e8-3569-4b44-93e6-ce397c7d49fe · outbound
WhisperFlow: speech foundation models in real time Review of speech-to-text recognition technology for enhancing learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ff1dba35-b561-490c-8c76-5ade1e72c3a1 · outbound
WhisperFlow: speech foundation models in real time Dissecting User-Perceived Latency of On-Device E2E Speech Recognition
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada30cfe-0565-4506-b4a1-d835bfc7a6b6 · outbound
WhisperFlow: speech foundation models in real time PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b63a376d-5d78-4531-94ae-2497a4566db1 · outbound
WhisperFlow: speech foundation models in real time Instantaneous Grammatical Error Correction with Shallow Aggressive Decoding
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73da8f7c-9092-44f6-8757-4afae33bd75c · outbound
WhisperFlow: speech foundation models in real time Intriguing properties of neural networks
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fff794c-ac5e-42d5-b128-04ba5146e1a8 · outbound
WhisperFlow: speech foundation models in real time Streaming trans- former asr with blockwise synchronous beam search
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 991113f4-2e3a-40c0-b0c3-38b0e68e9f90 · outbound
WhisperFlow: speech foundation models in real time The calo meeting speech recognition and understanding system
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3dc58165-3aab-40f6-a3b2-0a83fc81a30f · outbound
WhisperFlow: speech foundation models in real time Attention is all you need
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6fff2f0-5bbf-43a4-b963-9db803f49ac5 · outbound
WhisperFlow: speech foundation models in real time Simul-Whisper: Attention-Guided Streaming Whisper with Truncation Detection
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03387295-7265-473b-821a-24b96bc67d39 · outbound
WhisperFlow: speech foundation models in real time Turbocharge speech understanding with pilot inference
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03455b10-86cd-4b1a-ad0f-9a48a940dd5f · outbound
WhisperFlow: speech foundation models in real time ESPnet: End-to- end speech processing toolkit
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 42c84d03-9e8b-4e7b-a97f-b066a9915e7a · outbound
WhisperFlow: speech foundation models in real time Loongserve: Efficiently serving long-context large language models with elastic sequence parallelism
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f28cfe96-0ce5-4b57-be1b-aa55c22217d2 · outbound
WhisperFlow: speech foundation models in real time Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a5b0c0d-2af8-4d98-8484-c827f9c37b3f · outbound
WhisperFlow: speech foundation models in real time Toward human parity in con- versational speech recognition
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0f39021f-fe3e-452b-af6b-d3d1fd2dd17c · outbound
WhisperFlow: speech foundation models in real time Fast On-device LLM Inference with NPUs
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad35efc5-74ed-4041-9ed5-32b43527ef68 · outbound
WhisperFlow: speech foundation models in real time Inference with Reference: Lossless Acceleration of Large Language Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad388ebc-0738-4f51-a6b9-13e26e4f2cbe · outbound
WhisperFlow: speech foundation models in real time Transformer transducer: A streamable speech recog- nition model with transformer encoders and rnn-t loss
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25c14384-0ad6-4aa9-bafd-b20256e9b58d · outbound
WhisperFlow: speech foundation models in real time Black-box adversarial attacks on commercial speech platforms with minimal information
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 85c56d45-b835-4286-8d99-be546aa2d969 · outbound
WhisperFlow: speech foundation models in real time In 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI 24) , pages 193–210, 2024
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6a3efb1f-8404-416b-9c99-cd82ef5f4442 · outbound
WhisperFlow: speech foundation models in real time Unresolved cited work
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d164851d-0a5a-4698-92eb-9e96d9578ee8 · outbound
WhisperFlow: speech foundation models in real time doi:10.1145/3694715.3695948
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 818eced5-2f8e-4812-9425-a9adefa861aa · inbound
WhisperRT -- Turning Whisper into a Causal Streaming Model WhisperFlow: speech foundation models in real time
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4440a6c6-e16d-4a46-98e2-82eba009ed46 · inbound
Sink or SWIM: Tackling Real-Time ASR at Scale WhisperFlow: speech foundation models in real time
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.