VSRo-200 is the first large Romanian sentence-level lip-reading corpus; human labels beat pseudo-labels at matched size, but scaling pseudo-labels closes the gap and yields strong transfer to word recognition.
Training strategies for improved lip- reading,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
VSRo-200: A Romanian Visual Speech Recognition Dataset for Studying Supervision and Multimodal Robustness
VSRo-200 is the first large Romanian sentence-level lip-reading corpus; human labels beat pseudo-labels at matched size, but scaling pseudo-labels closes the gap and yields strong transfer to word recognition.