Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T05:01:56.219187Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 5 inbound Pith citation observations for arXiv:2502.03382.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T05:01:56.219187Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:28:52.222146Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T03:36:30.178589Z
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 368e8d34-37cf-4fbd-be4a-dd40b7593262 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f71584e-53a2-4c4c-ac65-e7da7007146c · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation A benchmark for evaluating machine translation metrics on dialects without standard orthography
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation db04918e-1588-4b62-94e2-e2ab7ccfb8f7 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Seamless: Multilingual Expressive and Streaming Speech Translation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4e9f196-c9d2-47b9-8c35-67858336164d · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Listen and Translate: A Proof of Concept for End-to-End Speech-to-Text Translation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9330bbfe-c1c0-4efa-ba2d-4a29429e849f · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Audiolm: A language modeling approach to audio generation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 34b01009-3b96-47c9-b488-b0d8b9ace1f1 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Wavlm: Large-scale self-supervised pre-training for full stack speech processing
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d89801ef-80e0-4155-a811-69c8263fe578 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Simple and controllable music generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4ddc70aa-6e8b-4598-b434-cfbfa288e96d · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Moshi: a speech-text foundation model for real-time dialogue
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7faa02a1-24d9-4656-abfc-2d13a9480396 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Daspeech: Directed acyclic transformer for fast and high-quality speech-to-speech translation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 08518d04-c4ad-44c0-8467-9002f0bddf52 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Gaussian Error Linear Units (GELUs)
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50749da1-2985-47e5-b8e5-8f0c8d5d4887 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation U nit Y : Two-pass direct speech-to-speech translation with discrete units
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b5976ed-61f7-4af7-a975-f8ebbc10f3ca · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation N., McNair, A
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 48cf3ee9-048c-46ad-8c88-689e1443a03f · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation J., Biadsy, F., Macherey, W., Johnson, M., Chen, Z., and Wu, Y
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9beecdd9-e835-4a2e-8a16-b6c0802f5206 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation T., Remez, T., and Pomerantz, R
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 05095930-e182-4e7f-962a-c527ea50e010 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation T., Wang, Q., and Zen, H
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e6ebff19-ee89-41c4-9881-4bc9efd3ca52 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation CTRL: A Conditional Transformer Language Model for Controllable Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24255d2b-6aa9-46a7-a490-196b1072a740 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation V., Buckley, C
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ce7b1ffc-5375-44e8-abcc-6c5082a82716 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Audiogen: Textually guided audio generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca2ee28e-bcc2-4b4c-b305-ca89eb803cc9 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation MADLAD-400: A multilingual and document-level large audited dataset
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9ec4525b-5246-484a-a059-e6eae5da9c5b · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Direct speech-to-speech translation with discrete units
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd3affd-56e0-4277-be05-ca9550fe94fc · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Autoregressive image generation using residual quantization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f92fafe5-50bb-47a0-9ed5-ab1b415a590b · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation and Hutter, F
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d1939de2-4210-4e52-a5ca-382d1616690a · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation whisper-timestamped
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 60b1f997-233b-4e5f-be6a-4472f42a01f8 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation J., Koehn, P., and Pino, J
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7891340f-a044-49e7-93d2-97500fbd21be · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation The ATR multilingual speech-to-speech translation system
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 57f85f41-24c6-476c-bf93-504b9269e077 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Over-Generation Cannot Be Rewarded: Length-Adaptive Average Lagging for Simultaneous Speech Translation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 903523d6-6300-4dcb-b1f3-22c82a8d28ff · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation A call for clarity in reporting BLEU scores
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6daa7a8-5227-4a5e-95ac-5cbefc82e74e · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation W., Xu, T., Brockman, G., McLeavey, C., and Sutskever, I
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6edda002-bfde-4916-ac13-a2239a5ad4e8 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Generating diverse high-fidelity images with vq-vae-2
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d378446-2f56-4326-839d-ba18092464e6 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation S imul S peech: End-to-end simultaneous speech to text translation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8156c230-6d5b-4196-81e1-391c6ab067b3 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation AudioPaLM: A Large Language Model That Can Speak and Listen
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e000f0b-d714-4bd1-9d26-2353563fc308 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation PySBD: Pragmatic Sentence Boundary Disambiguation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f51f16b1-a65a-4bca-8c9e-bfae812a60f2 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation GLU Variants Improve Transformer
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9fdb950-94ea-46d3-9e15-0baa1977ca41 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation N., and and, L
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3c3af593-c798-465e-942f-de4f1d60516f · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Verbmobil: Foundations of speech-to-speech translation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 91613f5d-e64e-475e-b9b5-78b1301e43e2 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Fairseq S 2 T : Fast speech-to-text modeling with fairseq
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28bb8ab5-b258-4122-b959-8943babad454 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation M., and Dupoux, E
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a21a9aa8-5682-4f4d-964e-55c2a8a53bf0 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Covost 2 and massively multilingual speech translation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcd0aecf-2797-4f6f-81af-9765b1c5006c · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation J., Chorowski, J., Jaitly, N., Wu, Y., and Chen, Z
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1ad2248-1f26-4077-b858-4f3daf0ffb90 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation UniAudio: An Audio Foundation Model Toward Universal Audio Generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf300c78-a469-47f9-a984-78049e551e54 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Soundstream: An end-to-end neural audio codec
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8cf7ae20-c6a1-44f5-900f-6a3bcafa2fe1 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Realtrans: End-to-end simultaneous speech translation with convolutional weighted-shrinking transformer
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 37dce851-5b88-42c3-824d-2a94cbe06235 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Streamspeech: Simultaneous speech-to-speech translation with multi-task learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdbce54b-8e80-4095-8595-534cee55da31 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Speechtokenizer: Unified speech tokenizer for speech language models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b55d94bf-86ef-4ebc-ab18-0c2c607ccbc7 · outbound
High-Fidelity Simultaneous Speech-To-Speech Translation Textless Streaming Speech-to-Speech Translation using Semantic Speech Tokens
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e1bcde44-36e0-4eca-ae94-34ec2c0fdf4f · inbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI High-Fidelity Simultaneous Speech-To-Speech Translation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9893ded1-943e-47cd-b054-faab046ad98c · inbound
Streaming Endpointer for Spoken Dialogue using Neural Audio Codecs and Label-Delayed Training High-Fidelity Simultaneous Speech-To-Speech Translation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 541cf954-a92b-44d0-ac2b-02138a4b836a · inbound
DiffSoundStream: Efficient Speech Tokenization via Diffusion Decoding High-Fidelity Simultaneous Speech-To-Speech Translation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8337e72c-d692-417b-90ce-912298b5ca36 · inbound
Regularized Entropy Information Adaptation with Temporal-Awareness Networks for Simultaneous Speech Translation High-Fidelity Simultaneous Speech-To-Speech Translation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 433a5b46-cec5-4904-b6a8-20cf5ea6305c · inbound
A Pocket Offline Model for Simultaneous Speech Translation as CUNI Submission to IWSLT 2026 High-Fidelity Simultaneous Speech-To-Speech Translation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.