A survey of 110 SimulST papers shows most systems rely on unrealistic human pre-segmented audio and inconsistent terminology, and it offers a taxonomy and recommendations to fix both.
In Proceedings of the 20th International Confer- ence on Spoken Language Translation (IWSLT 2023), pages 169–179, Toronto, Canada (in- person and online)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System?
A survey of 110 SimulST papers shows most systems rely on unrealistic human pre-segmented audio and inconsistent terminology, and it offers a taxonomy and recommendations to fix both.