VSRo-200 is the first large Romanian sentence-level lip-reading corpus; human labels beat pseudo-labels at matched size, but scaling pseudo-labels closes the gap and yields strong transfer to word recognition.
End-to-end lip reading in romanian with cross-lingual domain adaptation and lateral inhibition,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
VSRo-200: A Romanian Visual Speech Recognition Dataset for Studying Supervision and Multimodal Robustness
VSRo-200 is the first large Romanian sentence-level lip-reading corpus; human labels beat pseudo-labels at matched size, but scaling pseudo-labels closes the gap and yields strong transfer to word recognition.