Pith. sign in

REVIEW 2 cited by

Simultaneous Translation for Unsegmented Input: A Sliding Window Approach

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.09754 v1 pith:3U2B7VQE submitted 2022-10-18 cs.CL

classification cs.CL
keywords approachtranslationwindowonlineautomaticinputparallelsegmentation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In the cascaded approach to spoken language translation (SLT), the ASR output is typically punctuated and segmented into sentences before being passed to MT, since the latter is typically trained on written text. However, erroneous segmentation, due to poor sentence-final punctuation by the ASR system, leads to degradation in translation quality, especially in the simultaneous (online) setting where the input is continuously updated. To reduce the influence of automatic segmentation, we present a sliding window approach to translate raw ASR outputs (online or offline) without needing to rely on an automatic segmenter. We train translation models using parallel windows (instead of parallel sentences) extracted from the original training data. At test time, we translate at the window level and join the translated windows using a simple approach to generate the final translation. Experiments on English-to-German and English-to-Czech show that our approach improves 1.3--2.0 BLEU points over the usual ASR-segmenter pipeline, and the fixed-length window considerably reduces flicker compared to a baseline retranslation-based online SLT system.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Simulstream: Open-Source Toolkit for Evaluation and Demonstration of Streaming Speech-to-Text Translation Systems

    cs.CL 2025-12 conditional novelty 6.0 of 10

    Simulstream is the first open-source framework to jointly evaluate and demo streaming speech-to-text translation systems under both incremental and re-translation decoding, and its experiments suggest incremental Stre...

  2. MLLP-VRAIN UPV system for the IWSLT 2025 Simultaneous Speech Translation Translation task

    cs.CL 2025-06 conditional novelty 4.0 of 10

    A cascade of Whisper and NLLB, adapted with prefix training and streaming policies, achieves 29.8 BLEU on the IWSLT 2025 simultaneous speech translation test set.

Pith tools